Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Huggingface Trl GRPO Training
- Workflow:BerriAI Litellm Proxy Server Deployment
- Workflow:Microsoft Playwright AI agent driven testing
- Workflow:Kornia Kornia ONNX Model Pipeline
- Workflow:ARISE Initiative Robosuite Domain Randomization Training
- Workflow:Deepseek ai Janus Multimodal Understanding
- Workflow:DataTalksClub Data engineering zoomcamp dbt Analytics Transformation
- Workflow:Puppeteer Puppeteer Page Screenshot Capture
- Workflow:Hpcaitech ColossalAI Distributed GRPO Training
- Workflow:Intel Ipex llm RAG With LangChain
Principles
- Principle:Apache Hudi Compaction Need Assessment
- Principle:NVIDIA TransformerEngine Drop In LayerNorm Replacement
- Principle:DistrictDataLabs Yellowbrick Joint Plot Analysis
- Principle:Anthropics Anthropic sdk python Streaming Thinking Content
- Principle:Pyro ppl Pyro Global Configuration
- Principle:Turboderp org Exllamav2 Sampling Configuration
- Principle:Deepset ai Haystack Retrieval MRR Evaluation
- Principle:Neuml Txtai Search Explainability
- Principle:SeldonIO Seldon core Candidate Model Deployment
- Principle:Bitsandbytes foundation Bitsandbytes XPU SYCL Dequantization
Implementations
- Implementation:Interpretml Interpret PoissonDevianceRegressionObjective
- Implementation:Google deepmind Dm control Suite Cartpole
- Implementation:ThreeSR Awesome Inference Time Scaling GitHub Pull Request
- Implementation:Vespa engine Vespa ConfigSubscriber NextConfig
- Implementation:OpenGVLab InternVL Correctness Build Data
- Implementation:DevExpress Testcafe NativeAutomation Input
- Implementation:Scikit learn Scikit learn RocCurveDisplay
- Implementation:Apache Paimon BinaryRow
- Implementation:NVIDIA DALI Sphinx Config
- Implementation:Open compass VLMEvalKit VGRPBench Futoshiki
Heuristics
- Heuristic:OpenGVLab InternVL Gradient Checkpointing Memory
- Heuristic:Microsoft DeepSpeedExamples Gradient Checkpointing Tradeoff
- Heuristic:Vllm project Vllm KV Cache Block Size Selection
- Heuristic:Liu00222 Open Prompt Injection PPL Threshold Tuning
- Heuristic:Unslothai Unsloth Padding Free Packing
- Heuristic:DataExpert io Data engineer handbook SparkSession Singleton Pattern
- Heuristic:Microsoft LoRA Label Smoothing NLG
- Heuristic:Kubeflow Pipelines Component URL Commit SHA Pinning
- Heuristic:Datajuicer Data juicer Partition Size Tuning
- Heuristic:MaterializeInc Materialize CI Retry Strategies
Environments
- Environment:LLMBook zh LLMBook zh github io HuggingFace Transformers Stack
- Environment:Puppeteer Puppeteer Configuration Environment Variables
- Environment:Shiyu coder Kronos Comet ML Logging
- Environment:Getgauge Taiko Linux System Libraries
- Environment:ThreeSR Awesome Inference Time Scaling Python Runtime Environment
- Environment:Eventual Inc Daft AI Provider Dependencies
- Environment:DataExpert io Data engineer handbook PostgreSQL Docker Environment
- Environment:Helicone Helicone Cloudflare Workers Runtime
- Environment:Kubeflow Pipelines Python SDK
- Environment:Run llama Llama index Fsspec Remote Storage