Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Deepspeedai DeepSpeed ZeRO Distributed Training
- Workflow:Mlc ai Mlc llm Python Engine Inference
- Workflow:NVIDIA TransformerEngine Accelerate HF Llama With TE
- Workflow:Triton inference server Server LLM Deployment With TRT LLM
- Workflow:Onnx Onnx Model Validation
- Workflow:Recommenders team Recommenders Algorithm Benchmarking
- Workflow:Langfuse Langfuse Evaluation pipeline
- Workflow:Microsoft Playwright AI agent driven testing
- Workflow:Openai Whisper Word Level Timestamps
- Workflow:HKUDS AI Trader Agent Decision Loop
Principles
- Principle:Online ml River Online Preprocessing
- Principle:LLMBook zh LLMBook zh github io Deduplication
- Principle:Openai Openai python Chat Message Construction
- Principle:NVIDIA NeMo Aligner DPO Training
- Principle:Lm sys FastChat Condensed Rotary Embedding
- Principle:Langchain ai Langgraph Multi Agent Research Pattern
- Principle:Hiyouga LLaMA Factory Proximal Policy Optimization
- Principle:FMInference FlexLLMGen NVMe Disk Setup
- Principle:AUTOMATIC1111 Stable diffusion webui Application Lifecycle
- Principle:PrefectHQ Prefect DataFrame Transformation
Implementations
- Implementation:MaterializeInc Materialize CLI Workload Anonymize
- Implementation:Arize ai Phoenix Legacy Generate
- Implementation:Datahub project Datahub FieldPath Util
- Implementation:TobikoData Sqlmesh Context Invalidate Environment
- Implementation:Treeverse LakeFS Java SDK Model UsageReport
- Implementation:Infiniflow Ragflow RetrievalDocuments Component
- Implementation:LMCache LMCache Usage Context
- Implementation:Tensorflow Tfjs Model Save Test
- Implementation:EvolvingLMMs Lab Lmms eval MMSearch Plus Decrypt Utils
- Implementation:NVIDIA NeMo Curator FastText Filters
Heuristics
- Heuristic:DevExpress Testcafe Browser Connection Timeouts
- Heuristic:Deepset ai Haystack Document Splitting Defaults
- Heuristic:NVIDIA DALI Distributed Sharding Strategy
- Heuristic:Huggingface Alignment handbook Global Batch Size Scaling
- Heuristic:Rapidsai Cuml GPU Cache Alignment
- Heuristic:Apache Beam GC Thrashing Detection
- Heuristic:Google research Deduplicate text datasets Variable Width Pointer Optimization
- Heuristic:Ggml org Llama cpp Quantization Quality Tips
- Heuristic:Iamhankai Forest of Thought UCB Exploration Constant
- Heuristic:Groq Groq python Streaming Usage Stats
Environments
- Environment:Sgl project Sglang GitHub Actions
- Environment:TobikoData Sqlmesh Python Runtime
- Environment:Vibrantlabsai Ragas Optional NLP Metrics Environment
- Environment:Apache Dolphinscheduler Java Runtime
- Environment:Intel Ipex llm RAG LlamaIndex Environment
- Environment:HKUDS AI Trader Python LangChain Runtime
- Environment:Huggingface Transformers PyTorch 24 CUDA
- Environment:CarperAI Trlx Python Accelerate
- Environment:Huggingface Datatrove Inference GPU Environment
- Environment:Vllm project Vllm AWS ECR