Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Openai CLIP Linear probe evaluation
- Workflow:AUTOMATIC1111 Stable diffusion webui Textual inversion training
- Workflow:MarketSquare Robotframework browser Plugin Development
- Workflow:Datajuicer Data juicer Custom Operator Development
- Workflow:Huggingface Datatrove Minhash Deduplication
- Workflow:Apache Airflow Provider Distribution Development
- Workflow:Google deepmind Mujoco Model compilation and conversion
- Workflow:Teamcapybara Capybara Form Testing
- Workflow:Alibaba MNN Model Conversion Pipeline
- Workflow:Huggingface Diffusers Checkpoint Conversion
Principles
- Principle:Webdriverio Webdriverio Browser Session Creation
- Principle:SeldonIO Seldon core Usage Metrics Data Model
- Principle:Vespa engine Vespa Stemming
- Principle:Ggml org Llama cpp Sampling System
- Principle:Ggml org Llama cpp LoggingInfrastructure
- Principle:Haosulab ManiSkill CPU GPU Dual Backend
- Principle:Apache Druid Datasource Introspection
- Principle:Tencent Ncnn Neural Speech Synthesis
- Principle:Ggml org Llama cpp Partial JSON Healing
- Principle:Openai Openai agents python Stream Event Consumption
Implementations
- Implementation:TobikoData Sqlmesh LoadingContainer
- Implementation:Bentoml BentoML Bentoml Service Decorator
- Implementation:Huggingface Trl Get Dataset GRPO
- Implementation:Microsoft Autogen Component Schema Gen
- Implementation:Microsoft Onnxruntime FusedAdam FP16Optimizer
- Implementation:Open compass VLMEvalKit MEGABench Constrained Generation
- Implementation:Huggingface Diffusers Custom Init Isort
- Implementation:Webdriverio Webdriverio BrowserStack Types
- Implementation:Deepspeedai DeepSpeed AutoModel For Inference
- Implementation:Evidentlyai Evidently Legacy JSON Schema Match Feature
Heuristics
- Heuristic:Puppeteer Puppeteer Navigation Race Condition Avoidance
- Heuristic:Speechbrain Speechbrain Score Normalization Tips
- Heuristic:Spotify Luigi Dynamic Requirements Generator
- Heuristic:Allenai Open instruct GPU Memory Utilization
- Heuristic:EvolvingLMMs Lab Lmms eval Memory Cleanup After Inference
- Heuristic:Huggingface Peft Warning Deprecated Bone
- Heuristic:PeterL1n BackgroundMattingV2 ONNX Patch Method Compatibility
- Heuristic:Apache Shardingsphere Shadow Routing Hint First Fallback
- Heuristic:Google deepmind Mujoco Mesh Quality For Collision
- Heuristic:Mlc ai Web llm Low Resource Model Selection
Environments
- Environment:Speechbrain Speechbrain Speech Enhancement Dependencies
- Environment:Snorkel team Snorkel SpaCy NLP
- Environment:Datajuicer Data juicer Ray Cluster Environment
- Environment:EvolvingLMMs Lab Lmms eval API Credentials Environment
- Environment:Huggingface Transformers 3D Parallel Multi GPU
- Environment:Volcengine Verl Ray Distributed Environment
- Environment:ArroyoSystems Arroyo Rust Runtime
- Environment:Getgauge Taiko Linux System Libraries
- Environment:SeleniumHQ Selenium Contributor Development Environment
- Environment:EvolvingLMMs Lab Lmms eval Server Mode Environment