Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Datahub project Datahub Java SDK Metadata Emission
- Workflow:HKUDS AI Trader Multi Agent Comparison
- Workflow:Hiyouga LLaMA Factory LoRA SFT Finetuning
- Workflow:Kornia Kornia Edge Detection Pipeline
- Workflow:Vespa engine Vespa Linguistics text processing pipeline
- Workflow:Junyanz Pytorch CycleGAN and pix2pix Pretrained Inference
- Workflow:Microsoft Onnxruntime Distributed Model Training
- Workflow:Helicone Helicone Integrate Provider To Gateway
- Workflow:MaterializeInc Materialize Docker Image Build
- Workflow:Roboflow Rf detr Custom Dataset Finetuning
Principles
- Principle:Farama Foundation Gymnasium Video Recording
- Principle:CARLA simulator Carla Unreal Engine Build
- Principle:Huggingface Diffusers Prompt Encoding
- Principle:Dagster io Dagster Incremental Processing
- Principle:Pola rs Polars Input Data Preparation
- Principle:OpenRLHF OpenRLHF SFT Dataset Construction
- Principle:Pytorch Serve vLLM Model Configuration
- Principle:Interpretml Interpret Data Exploration
- Principle:Vllm project Vllm Speculative Engine Initialization
- Principle:Farama Foundation Gymnasium MuJoCo Locomotion
Implementations
- Implementation:Unstructured IO Unstructured BaseEmbeddingEncoder
- Implementation:Google deepmind Mujoco MJWarp Passive
- Implementation:InternLM Lmdeploy Daily Ete Test
- Implementation:Run llama Llama index MistralAIFinetuneEngine
- Implementation:ArroyoSystems Arroyo Confluent Connector
- Implementation:Infiniflow Ragflow RetrievalDocuments Component
- Implementation:Arize ai Phoenix Legacy Templates
- Implementation:Predibase Lorax Strategy Registry
- Implementation:Run llama Llama index BaseSelector
- Implementation:Mlc ai Mlc llm Text Streamer
Heuristics
- Heuristic:Volcengine Verl Sequence Length Balancing
- Heuristic:Ray project Ray Graceful Shutdown Timing
- Heuristic:Iamhankai Forest of Thought UCB Exploration Constant
- Heuristic:FMInference FlexLLMGen Pin Memory Tradeoffs
- Heuristic:Sdv dev SDV HMA Schema Simplification
- Heuristic:Sktime Pytorch forecasting Gradient Clipping Value
- Heuristic:ContextualAI HALOs LoRA Merge At Save
- Heuristic:EvolvingLMMs Lab Lmms eval Limit Flag Testing Only
- Heuristic:Dagster io Dagster Retry Strategy Configuration
- Heuristic:NVIDIA NeMo Aligner PPO Critic Warmup Tip
Environments
- Environment:Pola rs Polars Cloud Storage Environment
- Environment:Alibaba ROLL SGLang Inference Environment
- Environment:Huggingface Datasets JAX Integration
- Environment:SeldonIO Seldon core Python ML Dependencies Environment
- Environment:SeleniumHQ Selenium Selenium Manager Runtime
- Environment:Interpretml Interpret Blackbox Explainer Dependencies
- Environment:Eventual Inc Daft Ray Distributed Runner
- Environment:Intel Ipex llm RAG LangChain Environment
- Environment:Run llama Llama index OpenAI API Configuration
- Environment:Spotify Luigi Apache Spark