Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Treeverse LakeFS Write Audit Publish With Hooks
- Workflow:Turboderp org Exllamav2 Bulk Dataset Inference
- Workflow:Nightwatchjs Nightwatch Cucumber BDD Integration
- Workflow:Cleanlab Cleanlab Datalab Dataset Audit
- Workflow:Cleanlab Cleanlab Multiannotator Consensus
- Workflow:Huggingface Optimum Automatic Tensor Parallelization
- Workflow:Openai Openai agents python Multi Agent Handoff
- Workflow:Alibaba ROLL DPO Training Pipeline
- Workflow:Apache Kafka Topic Management
- Workflow:Cohere ai Cohere python Text Embedding
Principles
- Principle:Rapidsai Cuml Ranking Evaluation
- Principle:SqueezeAILab ETS Dependency Installation
- Principle:Langgenius Dify Workflow Execution Monitoring
- Principle:Huggingface Trl Reward Preference Dataset Loading
- Principle:Tencent Ncnn Foreign Model Conversion
- Principle:Pytorch Serve Accelerate Device Mapping
- Principle:Recommenders team Recommenders Data Loading MovieLens Pandas
- Principle:Pytorch Serve Speculative Decoding Inference
- Principle:Apache Kafka JIRA Issue Resolution
- Principle:Zai org CogVideo I2V Pipeline Loading
Implementations
- Implementation:ArroyoSystems Arroyo Kafka Source Tests
- Implementation:OpenHands OpenHands Integration Router Pattern
- Implementation:Langchain ai Langgraph UI Messages
- Implementation:Fastai Fastbook Export Load Learner
- Implementation:InternLM Lmdeploy BlockIterator
- Implementation:ArroyoSystems Arroyo Operator Trait
- Implementation:Lance format Lance Java FragmentMetadata
- Implementation:Princeton nlp SimPO Apply Chat Template
- Implementation:Explodinggradients Ragas RubricsScore Metric
- Implementation:Microsoft Autogen AssistantAgent Init Tools
Heuristics
- Heuristic:Fastai Fastbook Embedding Size Rule
- Heuristic:Apache Hudi Record Level Index Optimization
- Heuristic:Teamcapybara Capybara Animation Disabling For Tests
- Heuristic:Helicone Helicone Tiered Pricing Threshold Matching
- Heuristic:SqueezeAILab ETS Embedding Model GPU Collocation
- Heuristic:Haifengl Smile BFGS Convergence Tuning
- Heuristic:VainF Torch Pruning AutoGrad Dependency Graph
- Heuristic:Fede1024 Rust rdkafka Multi Version Dependency Hazard
- Heuristic:Huggingface Alignment handbook Sequence Packing Strategy
- Heuristic:Arize ai Phoenix Notebook Event Loop Patching
Environments
- Environment:Triton inference server Server GPU CUDA Runtime
- Environment:Iamhankai Forest of Thought OpenAI API Credentials
- Environment:Huggingface Datatrove Python Runtime
- Environment:DataExpert io Data engineer handbook PostgreSQL Docker Environment
- Environment:Interpretml Interpret Visualization Environment
- Environment:CARLA simulator Carla Python API Runtime
- Environment:DataTalksClub Data engineering zoomcamp PySpark Batch Environment
- Environment:DevExpress Testcafe Firefox Marionette
- Environment:Apache Spark Python Environment
- Environment:Unstructured IO Unstructured PDF Dependencies