Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Risingwavelabs Risingwave Docker Deployment
- Workflow:Online ml River Streaming Anomaly Detection
- Workflow:ClickHouse ClickHouse Building From Source
- Workflow:Mistralai Client python Streaming Chat Completion
- Workflow:Deepset ai Haystack Hybrid Document Search
- Workflow:NVIDIA NeMo Curator Fuzzy Deduplication
- Workflow:ARISE Initiative Robomimic Trained Policy Evaluation
- Workflow:Openclaw Openclaw Agent Message Loop
- Workflow:Hpcaitech ColossalAI LLaMA Continual Pretraining
- Workflow:Huggingface Datatrove FineWeb Dataset Creation
Principles
- Principle:Allenai Open instruct DPO Loss Dispatch
- Principle:PeterL1n BackgroundMattingV2 Dataset path configuration
- Principle:Neuml Txtai Semantic Query
- Principle:NVIDIA TransformerEngine Gemma Weight Loading
- Principle:Snorkel team Snorkel Transformation Function Definition
- Principle:Deepspeedai DeepSpeed CUDA Inference Primitives
- Principle:Mistralai Client python Finetuning Job Control
- Principle:Duckdb Duckdb Package Validation
- Principle:Spcl Graph of thoughts GoT Graph Topology Design
- Principle:Turboderp org Exllamav2 Batch Job Iteration
Implementations
- Implementation:TobikoData Sqlmesh NodePort
- Implementation:Lance format Lance Encoding Block Statistics
- Implementation:CARLA simulator Carla Python Map Bindings
- Implementation:Turboderp org Exllamav2 ExLlamaV2DynamicGenerator Init
- Implementation:Google deepmind Dm control MJCF Attribute
- Implementation:Isaac sim IsaacGymEnvs Serializable
- Implementation:HKUDS AI Trader Ainvoke With Retry
- Implementation:Microsoft Semantic kernel IFunctionInvocationFilter
- Implementation:Mit han lab Llm awq Serve Controller
- Implementation:Promptfoo Promptfoo Logger Browser
Heuristics
- Heuristic:Open compass VLMEvalKit WORLD SIZE Unset For Model Building
- Heuristic:Datahub project Datahub Gradle Formatting Over Direct Tools
- Heuristic:MarketSquare Robotframework browser Docker Chrome Security
- Heuristic:Spcl Graph of thoughts Budget Gated Benchmark Execution
- Heuristic:Huggingface Alignment handbook Liger Kernel Memory
- Heuristic:Triton inference server Server Model Instance Scaling
- Heuristic:Openai Openai python Fine Tuning Data Preparation Tips
- Heuristic:Duckdb Duckdb PR Submission Strategy
- Heuristic:Huggingface Alignment handbook DPO Beta Selection
- Heuristic:Groq Groq python Timeout Configuration
Environments
- Environment:Vllm project Vllm Environment Variables
- Environment:Marker Inc Korea AutoRAG API Keys Configuration
- Environment:Vllm project Vllm Benchmarks
- Environment:OpenRLHF OpenRLHF Flash Attention Environment
- Environment:Apache Flink Node Build Environment
- Environment:CrewAIInc CrewAI LLM Provider Credentials
- Environment:Huggingface Alignment handbook Python Transformers
- Environment:Microsoft DeepSpeedExamples SuperOffload Runtime
- Environment:Run llama Llama index Python LlamaIndex Core
- Environment:Huggingface Diffusers Training Environment