Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:CarperAI Trlx ILQL Offline Training
- Workflow:Webdriverio Webdriverio Cucumber BDD Testing
- Workflow:Volcengine Verl Multi Turn Tool Use Training
- Workflow:Openai Openai node Streaming To Client
- Workflow:CrewAIInc CrewAI Custom Tool Integration
- Workflow:Snorkel team Snorkel Slice Aware Training
- Workflow:CrewAIInc CrewAI Hierarchical Crew Execution
- Workflow:Apache Spark Release Process
- Workflow:Dotnet Machinelearning GenAI Causal LM Inference
- Workflow:Explodinggradients Ragas RAG Evaluation
Principles
- Principle:OpenRLHF OpenRLHF Reward Shaping
- Principle:Explodinggradients Ragas Prompt Persistence
- Principle:Langgenius Dify Vector Database Selection
- Principle:Mistralai Client python Chat Completion
- Principle:Allenai Open instruct SFT Data Processing
- Principle:Marker Inc Korea AutoRAG Schema Construction
- Principle:Facebookresearch Habitat lab Dataset and Scene Preparation
- Principle:Microsoft Playwright Test File Creation
- Principle:Pytorch Serve Metrics Monitoring
- Principle:Eventual Inc Daft Arrow FFI Compatibility
Implementations
- Implementation:Bentoml BentoML Cloud Login
- Implementation:Facebookresearch Habitat lab Dataset
- Implementation:Vllm project Vllm LLM Init
- Implementation:Microsoft Autogen Studio MCP Sidebar
- Implementation:Apache Druid ShowJsonOrStages
- Implementation:Promptfoo Promptfoo Plugins Data Catalog
- Implementation:SeleniumHQ Selenium Closure ListenerMap
- Implementation:Neuml Txtai LiteLLM Vectors
- Implementation:BerriAI Litellm Policy Templates Backup
- Implementation:Openai Openai python Response Custom Tool Call Output
Heuristics
- Heuristic:Predibase Lorax Quantization Backend Selection
- Heuristic:NVIDIA TransformerEngine Sequence Length Alignment
- Heuristic:Huggingface Datasets Flatten Indices Performance
- Heuristic:DistrictDataLabs Yellowbrick NaN Data Handling
- Heuristic:BerriAI Litellm Cooldown Threshold Tuning
- Heuristic:Deepspeedai DeepSpeed FP16 Convergence Tips
- Heuristic:Google deepmind Mujoco MJX Benchmarking Tips
- Heuristic:Junyanz Pytorch CycleGAN and pix2pix Identity Loss Color Preservation
- Heuristic:ArroyoSystems Arroyo Stateful Operator TTL
- Heuristic:Lance format Lance Vector Index Tuning
Environments
- Environment:Apache Spark Kubernetes Runtime
- Environment:Trailofbits Fickling PyTorch
- Environment:Scikit learn Scikit learn Python Runtime Environment
- Environment:AnswerDotAI RAGatouille GPU CUDA Runtime
- Environment:Togethercomputer Together python API Credentials
- Environment:Huggingface Datasets Python PyArrow Core
- Environment:Kserve Kserve Leader Worker Set
- Environment:OpenRLHF OpenRLHF vLLM Environment
- Environment:Apache Spark Release Build Environment
- Environment:Unstructured IO Unstructured Ingest CLI