Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Ggml org Llama cpp Model Perplexity Evaluation
- Workflow:Huggingface Alignment handbook ORPO Single Stage Alignment
- Workflow:CarperAI Trlx SFT Instruction Tuning
- Workflow:Huggingface Peft QLoRA SFT Finetuning
- Workflow:Kserve Kserve Deploying InferenceService
- Workflow:Ggml org Llama cpp Interactive Chat
- Workflow:PrefectHQ Prefect Web Scraping Pipeline
- Workflow:Nightwatchjs Nightwatch Component Testing
- Workflow:Huggingface Datatrove FineWeb Dataset Creation
- Workflow:Dagster io Dagster DSPy Optimization
Principles
- Principle:LaurentMazare Tch rs Pretrained Model Instantiation
- Principle:NVIDIA DALI TensorFlow Training Integration
- Principle:Teamcapybara Capybara Selector Modification
- Principle:Huggingface Datatrove Text Normalization Utilities
- Principle:Webdriverio Webdriverio WebDriver Bidi Protocol
- Principle:ARISE Initiative Robosuite Camera Projection Utilities
- Principle:Lm sys FastChat LoRA Adapter Saving
- Principle:Pola rs Polars Data Type Transformation
- Principle:Groq Groq python Batch Status Polling
- Principle:Bitsandbytes foundation Bitsandbytes Global INT8 Quantization
Implementations
- Implementation:ArroyoSystems Arroyo Running State
- Implementation:Datahub project Datahub MergeIntoCommandEdgeInputDatasetBuilder
- Implementation:Bentoml BentoML SSE Descriptor
- Implementation:Datahub project Datahub PendingMutationsException
- Implementation:LLMBook zh LLMBook zh github io Trainer Train LoRA
- Implementation:Speechbrain Speechbrain Train CommonLanguage LangId
- Implementation:Teamcapybara Capybara Spec Server
- Implementation:FlowiseAI Flowise ToolsListTable
- Implementation:Online ml River Stats Var
- Implementation:Sgl project Sglang Expert Specialization
Heuristics
- Heuristic:BerriAI Litellm Batch Size Flush Interval Tuning
- Heuristic:Allenai Open instruct Pre Init Torch Distributed
- Heuristic:Apache Beam Executor Shutdown Ordering
- Heuristic:MarketSquare Robotframework browser Windows Shell NPM Workaround
- Heuristic:Treeverse LakeFS Presigned URL Expiry Tip
- Heuristic:Haotian liu LLaVA Gradient Checkpointing Memory Optimization
- Heuristic:AnswerDotAI RAGatouille Index Rebuild Vs Update Decision
- Heuristic:Huggingface Trl DeepSpeed ZeRO3 Generation Tradeoff
- Heuristic:Lucidrains X transformers Rotary Position Embedding Selection
- Heuristic:InternLM Lmdeploy KV Cache Memory Tuning
Environments
- Environment:Marker Inc Korea AutoRAG API Keys Configuration
- Environment:NVIDIA NeMo Aligner TensorRT LLM Acceleration Environment
- Environment:OpenHands OpenHands SaaS Server Environment
- Environment:Fastai Fastbook Jupyter Notebook Environment
- Environment:DataTalksClub Data engineering zoomcamp Kafka Confluent Environment
- Environment:Guardrails ai Guardrails OpenTelemetry Tracing
- Environment:CARLA simulator Carla Simulation Runtime
- Environment:Tensorflow Serving Docker Runtime Environment
- Environment:Huggingface Optimum Python Core Dependencies
- Environment:Vespa engine Vespa POSIX Mmap Log Control