Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Eventual Inc Daft Distributed UDF Processing
- Workflow:LMCache LMCache KV Cache Offloading
- Workflow:NVIDIA TransformerEngine FSDP Distributed Training
- Workflow:Microsoft DeepSpeedExamples CIFAR10 Getting Started
- Workflow:Trailofbits Fickling Safe ML Model Loading
- Workflow:NVIDIA NeMo Curator Text Curation Pipeline
- Workflow:EvolvingLMMs Lab Lmms eval Custom Task Creation
- Workflow:Pyro ppl Pyro SVI Training
- Workflow:Treeverse LakeFS Write Audit Publish With Hooks
- Workflow:TobikoData Sqlmesh Model development and testing
Principles
- Principle:Spotify Luigi Metrics Collection
- Principle:Mlc ai Web llm Web Worker Engine Proxy
- Principle:Haotian liu LLaVA CLI Interactive Chat
- Principle:Huggingface Datatrove Media Reading Framework
- Principle:Haosulab ManiSkill Trajectory Replay Conversion
- Principle:Guardrails ai Guardrails Stream Invocation
- Principle:Fastai Fastbook Backpropagation
- Principle:Bentoml BentoML Cloud Endpoint Invocation
- Principle:Dotnet Machinelearning Feature Engineering
- Principle:Triton inference server Server Config Optimization
Implementations
- Implementation:Arize ai Phoenix Datasets Create Dataset
- Implementation:Infiniflow Ragflow UseRenameDataset Hook
- Implementation:NVIDIA TransformerEngine Cpp GEMM
- Implementation:Google deepmind Dm control CMU Mocap Initializer
- Implementation:Evidentlyai Evidently Legacy Runner
- Implementation:Scikit learn Scikit learn VarianceThreshold
- Implementation:Deepset ai Haystack SentenceTransformersTextEmbedder
- Implementation:Hiyouga LLaMA Factory Logging
- Implementation:Ggml org Llama cpp Peg Parser
- Implementation:Neuml Txtai TokenDetection
Heuristics
- Heuristic:Huggingface Trl Disable Dropout For RL Training
- Heuristic:Open compass VLMEvalKit Video Frame Sampling Configuration
- Heuristic:Scikit learn Scikit learn Warning Deprecated PassiveAggressive
- Heuristic:Heibaiying BigData Notes Kafka Consumer Offset Strategy Tip
- Heuristic:Scikit learn contrib Imbalanced learn KNeighbors Selection Tips
- Heuristic:Openai Evals Thread Tuning
- Heuristic:OpenGVLab InternVL Gradient Checkpointing Memory
- Heuristic:Hiyouga LLaMA Factory Gradient Checkpointing Memory Optimization
- Heuristic:MarketSquare Robotframework browser MacOS Sonoma Startup Delay
- Heuristic:Risingwavelabs Risingwave Docker Memory Allocation
Environments
- Environment:Eric mitchell Direct preference optimization PyTorch CUDA
- Environment:BerriAI Litellm Redis Cache Backend
- Environment:Alibaba ROLL vLLM Inference Environment
- Environment:Ray project Ray Python Runtime Environment
- Environment:Haosulab ManiSkill GPU CUDA Simulation
- Environment:Speechbrain Speechbrain PyTorch CUDA Runtime
- Environment:Openai Openai agents python Voice Dependencies
- Environment:Openai Whisper Triton
- Environment:Huggingface Datatrove Inference GPU Environment
- Environment:Recommenders team Recommenders GPU CUDA Environment