Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:CarperAI Trlx SFT Instruction Tuning
- Workflow:ClickHouse ClickHouse Contributing Pull Request
- Workflow:VainF Torch Pruning LLM Structural Pruning
- Workflow:Ggml org Llama cpp Model Perplexity Evaluation
- Workflow:NVIDIA TransformerEngine Accelerate HF Llama With TE
- Workflow:Bitsandbytes foundation Bitsandbytes 4bit QLoRA Inference
- Workflow:SeleniumHQ Selenium Selenium Grid Deployment
- Workflow:Langfuse Langfuse Prompt management lifecycle
- Workflow:Apache Kafka Coordinator Runtime Lifecycle
- Workflow:Getgauge Taiko Form Interaction Testing
Principles
- Principle:Online ml River Estimator Validation Framework
- Principle:Alibaba MNN Diffusion MNN Conversion
- Principle:BerriAI Litellm Integration Selection
- Principle:Volcengine Verl Policy Loss Optimization
- Principle:Pytorch Serve Inference Handler Development
- Principle:Apache Paimon FAISS Index Configuration
- Principle:Googleapis Python genai History Management
- Principle:Open compass VLMEvalKit Video Inference Orchestration
- Principle:Google deepmind Mujoco Binary Model Loading
- Principle:Facebookresearch Audiocraft Latent Decoding and Audio Output
Implementations
- Implementation:Huggingface Open r1 Run Benchmark Jobs
- Implementation:Haotian liu LLaVA Load Pretrained Model
- Implementation:OpenHands OpenHands ResendKeycloakSync
- Implementation:Farama Foundation Gymnasium EzPickle
- Implementation:Rapidsai Cuml MulticlassClassifiers
- Implementation:Online ml River FeatureSelection SelectKBest
- Implementation:NVIDIA NeMo Curator Wikipedia URLGenerator
- Implementation:Apache Paimon AuthFactory
- Implementation:Huggingface Datatrove LineStats
- Implementation:TobikoData Sqlmesh MessageContainer
Heuristics
- Heuristic:LMCache LMCache Health Monitor Thresholds
- Heuristic:EvolvingLMMs Lab Lmms eval Memory Cleanup After Inference
- Heuristic:PeterL1n BackgroundMattingV2 Training Batch Size And Resolution
- Heuristic:Truera Trulens Rate Limiting And Retry Strategy
- Heuristic:Google research Deduplicate text datasets Ulimit File Descriptors For Merge
- Heuristic:Farama Foundation Gymnasium Render Mode Selection Guide
- Heuristic:Haifengl Smile Quarkus Async Context Handling
- Heuristic:Langchain ai Langchain Deprecation Version Tracking
- Heuristic:Danijar Dreamerv3 Adaptive Gradient Clipping
- Heuristic:LLMBook zh LLMBook zh github io Deduplication Ngram Threshold
Environments
- Environment:Mlc ai Mlc llm CUDA GPU Environment
- Environment:Cleanlab Cleanlab Datalab Dependencies
- Environment:Ray project Ray CI Build Matrix Environment
- Environment:Mlfoundations Open flamingo HuggingFace Open CLIP Dependencies
- Environment:Eric mitchell Direct preference optimization PyTorch CUDA
- Environment:Microsoft BIPIA Python CUDA GPU Environment
- Environment:Vllm project Vllm CUDA Runtime
- Environment:DataTalksClub Data engineering zoomcamp Docker PostgreSQL Python Environment
- Environment:Spotify Luigi AWS S3 Storage
- Environment:OpenRLHF OpenRLHF Ray Distributed Environment