Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:TobikoData Sqlmesh Github CICD automation
- Workflow:Fastai Fastbook Tabular Modeling
- Workflow:Princeton nlp SimPO SimPO Training
- Workflow:Huggingface Transformers Model Benchmarking
- Workflow:Ggml org Ggml Vision Model Inference
- Workflow:Iamhankai Forest of Thought CGDM Post Processing
- Workflow:Mlc ai Mlc llm Mobile Deployment
- Workflow:Hiyouga LLaMA Factory Full Parameter SFT
- Workflow:ARISE Initiative Robosuite Teleoperation
- Workflow:Facebookresearch Audiocraft JASCO Conditioned Music Generation
Principles
- Principle:PeterL1n BackgroundMattingV2 Checkpoint management
- Principle:Tensorflow Serving Executor Abstraction
- Principle:LLMBook zh LLMBook zh github io Deduplication
- Principle:Promptfoo Promptfoo Configuration Rendering
- Principle:Hiyouga LLaMA Factory Chat Template System
- Principle:Tencent Ncnn Non Maximum Suppression
- Principle:Puppeteer Puppeteer Accessibility Tree Inspection
- Principle:Datahub project Datahub Scheduled Ingestion
- Principle:Volcengine Verl GAE Advantage Estimation
- Principle:Run llama Llama index Embedding Finetune Execution
Implementations
- Implementation:Microsoft BIPIA HF Trainer For Defense
- Implementation:Openai Openai node Realtime Client Events
- Implementation:ArroyoSystems Arroyo Nats Source
- Implementation:Hiyouga LLaMA Factory Misc Utils
- Implementation:Hpcaitech ColossalAI Launch From Torch
- Implementation:ArroyoSystems Arroyo Sink V2 Migration
- Implementation:Run llama Llama index RecursiveRetriever
- Implementation:Bentoml BentoML InferenceAPI
- Implementation:Alibaba ROLL SFTWorker Train Step
- Implementation:Datajuicer Data juicer CleanHtmlMapper
Heuristics
- Heuristic:Mlflow Mlflow Nested Run Organization
- Heuristic:Apache Kafka Container JMX RMI Port Tip
- Heuristic:Scikit learn contrib Imbalanced learn Sampling Strategy Selection
- Heuristic:Langchain ai Langgraph Checkpointer Selection Guide
- Heuristic:FMInference FlexLLMGen Pin Memory Tradeoffs
- Heuristic:Allenai Open instruct Pre Init Torch Distributed
- Heuristic:Mlflow Mlflow Async Logging Best Practices
- Heuristic:Microsoft BIPIA LLAMA Pad Token Workaround
- Heuristic:Apache Airflow DAG Top Level Code Avoidance
- Heuristic:Isaac sim IsaacGymEnvs JIT Profiling Optimization
Environments
- Environment:Sgl project Sglang GitHub Actions
- Environment:Arize ai Phoenix Phoenix Server Runtime
- Environment:Sdv dev SDV GPU CUDA Support
- Environment:Mlc ai Web llm Chrome Extension Manifest V3
- Environment:Rapidsai Cuml Dask Distributed
- Environment:Kserve Kserve Cert Manager
- Environment:TobikoData Sqlmesh Web UI Stack
- Environment:Mlc ai Mlc llm Metal macOS iOS Environment
- Environment:Bigscience workshop Petals CUDA Server
- Environment:Pola rs Polars Python Runtime Environment