Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Recommenders team Recommenders Neural Collaborative Filtering
- Workflow:Danijar Dreamerv3 Evaluation Only
- Workflow:FlagOpen FlagEmbedding Embedder Finetuning
- Workflow:Spotify Luigi Central Scheduler Deployment
- Workflow:Apache Druid SQL Query Execution
- Workflow:Liu00222 Open Prompt Injection DataSentinel Detection
- Workflow:Risingwavelabs Risingwave Docker Deployment
- Workflow:Kornia Kornia Edge Detection Pipeline
- Workflow:Zai org CogVideo Diffusers Image to Video Inference
- Workflow:OpenRLHF OpenRLHF Knowledge Distillation
Principles
- Principle:Huggingface Datatrove Media Filtering Framework
- Principle:BerriAI Litellm Fine Tuned Model Usage
- Principle:Shiyu coder Kronos Qlib Data Preprocessing
- Principle:Google deepmind Dm control Multi Agent Episode Loop
- Principle:Norrrrrrr lyn WAInjectBench LoRA Adapter Injection
- Principle:Interpretml Interpret Rule Based Classification
- Principle:Huggingface Transformers Distributed Checkpointing
- Principle:Lance format Lance Approximate Nearest Neighbor Search
- Principle:Avhz RustQuant Bond Pricing
- Principle:Kubeflow Kubeflow Tune Hyperparameters
Implementations
- Implementation:Duckdb Duckdb Brotli Transform
- Implementation:FlagOpen FlagEmbedding LLM Embedder Eval Retrieval
- Implementation:Google deepmind Mujoco Engine Passive
- Implementation:Truera Trulens Selector Init
- Implementation:Danijar Dreamerv3 Logger And Report
- Implementation:Allenai Open instruct TensorCache
- Implementation:Pyro ppl Pyro TraceEnum ELBO Loss
- Implementation:Hpcaitech ColossalAI Load QA Chain
- Implementation:Eventual Inc Daft AI Embed Text
- Implementation:Langfuse Langfuse Analytics Integration Types
Heuristics
- Heuristic:Eric mitchell Direct preference optimization FSDP Batch Size Per GPU
- Heuristic:Google research Deduplicate text datasets Ulimit File Descriptors For Merge
- Heuristic:Princeton nlp SimPO Hyperparameter Tuning
- Heuristic:Fastai Fastbook Progressive Resizing
- Heuristic:OpenHands OpenHands Streamable HTTP Over SSE
- Heuristic:Langfuse Langfuse LLM Rate Limit 24h Abandon
- Heuristic:EvolvingLMMs Lab Lmms eval Memory Cleanup After Inference
- Heuristic:Onnx Onnx Big Endian Byte Order Handling
- Heuristic:Treeverse LakeFS Retry Backoff Configuration
- Heuristic:LMCache LMCache Memory Fragmentation Budget
Environments
- Environment:Iterative Dvc Git SCM Environment
- Environment:EvolvingLMMs Lab Lmms eval API Credentials Environment
- Environment:Huggingface Diffusers Training Environment
- Environment:Pyro ppl Pyro CUDA GPU Acceleration
- Environment:CarperAI Trlx DeepSpeed Multi GPU
- Environment:Mlfoundations Open flamingo Evaluation Dependencies
- Environment:Allenai Open instruct Python 3 12 Runtime
- Environment:Huggingface Transformers Flash Attention 2 Env
- Environment:Pytorch Serve vLLM Engine Environment
- Environment:Huggingface Datasets PyTorch Integration