Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Deepset ai Haystack Hybrid Document Search
- Workflow:Googleapis Python genai Model Fine Tuning
- Workflow:Vllm project Vllm Vision Language Inference
- Workflow:Princeton nlp Tree of thought llm ToT BFS experiment
- Workflow:Huggingface Open r1 Model Evaluation
- Workflow:Pyro ppl Pyro VAE Training
- Workflow:Recommenders team Recommenders SAR Collaborative Filtering
- Workflow:FMInference FlexLLMGen HELM Benchmark Evaluation
- Workflow:Googleapis Python genai Context Caching
- Workflow:Haotian liu LLaVA Two Stage Pretraining and Finetuning
Principles
- Principle:Pytorch Serve Model Artifact Configuration
- Principle:Apache Kafka Remote Push
- Principle:Cleanlab Cleanlab Multilabel Label Quality Scoring
- Principle:Google deepmind Mujoco Fluid Interaction
- Principle:Intel Ipex llm LoRA Adapter Injection
- Principle:Huggingface Diffusers Video Denoising
- Principle:Ucbepic Docetl Pipeline Optimization Search
- Principle:Apache Druid Schema Discovery
- Principle:DataTalksClub Data engineering zoomcamp Environment Setup
- Principle:Eric mitchell Direct preference optimization Batch Data Pipeline
Implementations
- Implementation:BerriAI Litellm Vector Stores API
- Implementation:MaterializeInc Materialize Cut Release Verify
- Implementation:CrewAIInc CrewAI Knowledge Constructor
- Implementation:Nautechsystems Nautilus trader TestInstrumentProvider Factory
- Implementation:Apache Flink WritableSerializer
- Implementation:ClickHouse ClickHouse Clickhouse Client Interactive
- Implementation:OpenHands OpenHands Dialog
- Implementation:BerriAI Litellm Check Responses Cost
- Implementation:Treeverse LakeFS S3 Path Convention
- Implementation:Microsoft DeepSpeedExamples Vision Transformer Model
Heuristics
- Heuristic:Tensorflow Serving Warning Deprecated CreateTfrtSavedModel Raw
- Heuristic:Hpcaitech ColossalAI Gradient Checkpointing Memory Tip
- Heuristic:Deepset ai Haystack Pipeline Max Runs Safety Limit
- Heuristic:Huggingface Alignment handbook Global Batch Size Scaling
- Heuristic:Apache Paimon Compression Tuning
- Heuristic:Gretelai Gretel synthetics Parallel Generation CUDA Disable
- Heuristic:Alibaba ROLL KL Coefficient Tuning
- Heuristic:Avhz RustQuant Learning Rate Tuning
- Heuristic:OpenGVLab InternVL Packed Training Buffer Management
- Heuristic:Hpcaitech ColossalAI Flash Attention Dtype Restriction
Environments
- Environment:LMCache LMCache VLLM Serving Engine
- Environment:Unstructured IO Unstructured Profiling Tools
- Environment:Facebookresearch Audiocraft FAD TensorFlow Environment
- Environment:Microsoft Onnxruntime Distributed Training Environment
- Environment:Microsoft Semantic kernel ONNX CUDA Environment
- Environment:Online ml River Build Toolchain
- Environment:Microsoft Autogen Studio Server Environment
- Environment:Predibase Lorax CUDA GPU Runtime
- Environment:FlowiseAI Flowise Node Runtime Environment
- Environment:Hpcaitech ColossalAI ColossalChat Training Environment