Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Getgauge Taiko Interactive Test Recording
- Workflow:Microsoft Onnxruntime Python Inference Pipeline
- Workflow:Google deepmind Mujoco Model compilation and conversion
- Workflow:Onnx Onnx Model Creation
- Workflow:Puppeteer Puppeteer Page Screenshot Capture
- Workflow:Deepspeedai DeepSpeed ZeRO Distributed Training
- Workflow:SeldonIO Seldon core AB Testing Experiment
- Workflow:Astronomer Astronomer cosmos Kubernetes dbt execution
- Workflow:Huggingface Datatrove Common Crawl Processing
- Workflow:Unstructured IO Unstructured Connector Ingest Pipeline
Principles
- Principle:Huggingface Optimum Task Specific Preprocessing
- Principle:DataTalksClub Data engineering zoomcamp GCS Data Upload
- Principle:Openclaw Openclaw Version Update
- Principle:Intel Ipex llm Data Preparation For Finetuning
- Principle:Apache Flink File Source Framework
- Principle:Open compass VLMEvalKit MCQ Prompt Construction
- Principle:Ggml org Llama cpp Server Build
- Principle:Apache Spark Release Version Tagging
- Principle:Ollama Ollama GGUF Model Conversion Qwen25Vl
- Principle:Pyro ppl Pyro Time Series Forecasting
Implementations
- Implementation:Neuml Txtai RDBMS Graph
- Implementation:Bentoml BentoML Gradio Mount
- Implementation:Puppeteer Puppeteer Cdp Frame
- Implementation:Mlc ai Mlc llm MLC Chat Config
- Implementation:Langfuse Langfuse PromptService Cache
- Implementation:TobikoData Sqlmesh Model Artifact
- Implementation:BerriAI Litellm Audit Logging Endpoints
- Implementation:Ray project Ray Repro CI Tool
- Implementation:CARLA simulator Carla VehiclePIDController Run Step
- Implementation:ArroyoSystems Arroyo Udaf Wrapper
Heuristics
- Heuristic:Apache Dolphinscheduler HikariCP Pool Tuning
- Heuristic:DevExpress Testcafe Assertion Retry Timing
- Heuristic:Facebookresearch Habitat lab Warning Deprecated Legacy UI System
- Heuristic:Facebookresearch Audiocraft Codebook Dead Code Expiration
- Heuristic:AUTOMATIC1111 Stable diffusion webui NaN Detection And Precision Fixes
- Heuristic:LMCache LMCache Memory Fragmentation Budget
- Heuristic:Trailofbits Fickling Severity Threshold Selection
- Heuristic:Mbzuai oryx Awesome LLM Post training Reference Citation Cap 200
- Heuristic:Unstructured IO Unstructured Strategy Fallback Chain
- Heuristic:Ucbepic Docetl Reduce Parallel Fold Tuning
Environments
- Environment:Puppeteer Puppeteer Node 18 Runtime
- Environment:Huggingface Trl PEFT LoRA Environment
- Environment:ChenghaoMou Text dedup Suffix Array External Tools
- Environment:Apache Shardingsphere Java Runtime Environment
- Environment:Haosulab ManiSkill Python SAPIEN Core
- Environment:Protectai Modelscan Python Core Runtime
- Environment:Promptfoo Promptfoo Provider API Keys
- Environment:DataTalksClub Data engineering zoomcamp Dbt DuckDB Environment
- Environment:Bigscience workshop Petals CUDA Server
- Environment:Run llama Llama index OpenAI API Configuration