Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Apache Flink File Source Pipeline
- Workflow:Huggingface Trl Direct Preference Optimization
- Workflow:Huggingface Open r1 SFT Distillation
- Workflow:Lakeraai Pint benchmark Custom Dataset Benchmarking
- Workflow:Mlc ai Web llm Chrome Extension Integration
- Workflow:Alibaba MNN Model Compression
- Workflow:Helicone Helicone LLM Response Normalization
- Workflow:Deepset ai Haystack Hybrid Document Search
- Workflow:PeterL1n BackgroundMattingV2 Video matting inference
- Workflow:DataExpert io Data engineer handbook PySpark Job Testing
Principles
- Principle:Apache Hudi Clustering Layout Analysis
- Principle:OpenRLHF OpenRLHF PPO Policy Loss
- Principle:Allenai Open instruct Causal LM Loading
- Principle:Openclaw Openclaw Browser CDP Relay
- Principle:Liu00222 Open Prompt Injection Model Creation
- Principle:Puppeteer Puppeteer Browser Binary Installation
- Principle:Lance format Lance Vector Index Building
- Principle:DataExpert io Data engineer handbook Kafka Source Table Definition
- Principle:Togethercomputer Together python Batch Job Creation
- Principle:CARLA simulator Carla Synchronous Simulation Mode
Implementations
- Implementation:Ray project Ray Ray Shutdown
- Implementation:Vllm project Vllm SGL GEMM INT8
- Implementation:Microsoft Autogen MagenticOne Prompts
- Implementation:Apache Paimon RenamingSnapshotCommit
- Implementation:OpenRLHF OpenRLHF Get llm for sequence regression
- Implementation:Ggml org Llama cpp GGUF Example
- Implementation:Teamcapybara Capybara RSpec Matchers HaveSelector
- Implementation:Iterative Dvc Update Meta
- Implementation:Turboderp org Exllamav2 Ext Cache
- Implementation:OpenGVLab InternVL CustomLayerDecayOptimizerConstructor
Heuristics
- Heuristic:Bentoml BentoML Warning Deprecated Server Module
- Heuristic:AnswerDotAI RAGatouille Index Rebuild Vs Update Decision
- Heuristic:Togethercomputer Together python Repetition Penalty Conflict
- Heuristic:LLMBook zh LLMBook zh github io LoRA Initialization Strategy
- Heuristic:Huggingface Optimum Device Offload Constraints
- Heuristic:ARISE Initiative Robosuite Domain Randomization Tuning
- Heuristic:Explodinggradients Ragas LLM Temperature Defaults
- Heuristic:Microsoft Autogen Warning Deprecated JSON Env Files
- Heuristic:Vespa engine Vespa Config Polling Timeout Tuning
- Heuristic:Romsto Speculative Decoding Gamma Tuning
Environments
- Environment:Huggingface Datasets SQL Dependencies
- Environment:Apache Flink Java Build Environment
- Environment:MaterializeInc Materialize Dbt Materialize Runtime
- Environment:Lakeraai Pint benchmark Python 310 With Transformers
- Environment:AUTOMATIC1111 Stable diffusion webui GPU Compute Backend
- Environment:Langgenius Dify Python Backend Environment
- Environment:Openai Whisper Numba
- Environment:Deepset ai Haystack HuggingFace Model Environment
- Environment:Huggingface Transformers 3D Parallel Multi GPU
- Environment:Marker Inc Korea AutoRAG Japanese NLP Dependencies