Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:CARLA simulator Carla Sensor Data Collection
- Workflow:Cohere ai Cohere python Model Finetuning
- Workflow:Mlc ai Web llm Web Worker Deployment
- Workflow:Dagster io Dagster DSPy Optimization
- Workflow:Langgenius Dify App Creation and Configuration
- Workflow:Facebookresearch Audiocraft JASCO Conditioned Music Generation
- Workflow:Teamcapybara Capybara Element Finding And Interaction
- Workflow:Huggingface Optimum Model Export
- Workflow:Snorkel team Snorkel Weak Supervision Pipeline
- Workflow:Mlc ai Web llm Structured Output Generation
Principles
- Principle:Tensorflow Serving Remote Predict Op
- Principle:Langchain ai Langgraph Thread Lifecycle Management
- Principle:Ucbepic Docetl Pipeline Optimization Search
- Principle:Online ml River Tree Split Strategies
- Principle:Astronomer Astronomer cosmos Graph Parsing and Task Generation
- Principle:Huggingface Datatrove Local Pipeline Execution
- Principle:Protectai Llm guard Input Scanner Factory Pattern
- Principle:Datahub project Datahub Deployment Verification
- Principle:Facebookresearch Habitat lab HITL Environment Setup
- Principle:Arize ai Phoenix LLM Provider Configuration
Implementations
- Implementation:CARLA simulator Carla Buffer
- Implementation:MaterializeInc Materialize CI Closed Issues Detect
- Implementation:DevExpress Testcafe TestRunTracker
- Implementation:Elevenlabs Elevenlabs python ChapterResponse
- Implementation:Run llama Llama index DocstoreStrategy Configuration
- Implementation:Open compass VLMEvalKit EgoExoBench Utils
- Implementation:Apache Airflow Timetable Protocol
- Implementation:Microsoft Playwright Server Android
- Implementation:Ucbepic Docetl ComponentUtils
- Implementation:Neuml Txtai SparseVectors Base
Heuristics
- Heuristic:Unslothai Unsloth LoRA Rank Selection
- Heuristic:Roboflow Rf detr Small Dataset Oversampling
- Heuristic:Diagram of thought Diagram of thought Strict Vs Flexible Critic Rigor
- Heuristic:Neuml Txtai MacOS Stability Workarounds
- Heuristic:Obss Sahi Overlap Ratio Selection
- Heuristic:Scikit learn contrib Imbalanced learn Sampling Strategy Selection
- Heuristic:LLMBook zh LLMBook zh github io LoRA Initialization Strategy
- Heuristic:Alibaba MNN Backend Selection Guide
- Heuristic:MaterializeInc Materialize Docker Image Cache Lookup
- Heuristic:Langgenius Dify SQL Escape Backslash First
Environments
- Environment:Datahub project Datahub Python Ingestion
- Environment:Infiniflow Ragflow GPU CUDA Environment
- Environment:InternLM Lmdeploy CUDA GPU Runtime
- Environment:PacktPublishing LLM Engineers Handbook API Credentials
- Environment:FlagOpen FlagEmbedding GPU Accelerator Environment
- Environment:Trailofbits Fickling Python Runtime
- Environment:Gretelai Gretel synthetics TensorFlow GPU Environment
- Environment:LLMBook zh LLMBook zh github io PyTorch CUDA GPU Environment
- Environment:Microsoft DeepSpeedExamples RLHF Training Environment
- Environment:Langchain ai Langgraph Python Runtime Environment