Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:InternLM Lmdeploy VLM Inference Pipeline
- Workflow:Speechbrain Speechbrain Speech Separation Training
- Workflow:Tensorflow Serving Kubernetes Deployment
- Workflow:Heibaiying BigData Notes Storm Topology Development
- Workflow:Truera Trulens RAG Evaluation With LangChain
- Workflow:Cypress io Cypress Local Development Environment
- Workflow:Apache Hudi Flink Table Clustering
- Workflow:CarperAI Trlx SFT Instruction Tuning
- Workflow:Huggingface Diffusers ControlNet Guided Generation
- Workflow:Microsoft Onnxruntime ORTModule Training
Principles
- Principle:Cypress io Cypress DOM Assertion and Interaction
- Principle:Ggml org Llama cpp Logging System
- Principle:Online ml River Page Hinkley Drift Detection
- Principle:Apache Dolphinscheduler Synchronous RPC Execution
- Principle:EvolvingLMMs Lab Lmms eval PassAtK Evaluation
- Principle:Isaac sim IsaacGymEnvs Environment Setup
- Principle:Langgenius Dify BrandingConfiguration
- Principle:Sgl project Sglang Server Arguments Configuration
- Principle:Langchain ai Langgraph SDK Error Handling
- Principle:Nautechsystems Nautilus trader Strategy Shutdown
Implementations
- Implementation:OpenGVLab InternVL Segmentation Test
- Implementation:SeleniumHQ Selenium HasCasting
- Implementation:Tencent Ncnn Darknet2ncnn
- Implementation:Astronomer Astronomer cosmos TrinoLDAPProfileMapping
- Implementation:Speechbrain Speechbrain LibriSpeech LM Dataset
- Implementation:Guardrails ai Guardrails Validator Validate
- Implementation:Datajuicer Data juicer FixUnicodeMapper
- Implementation:FMInference FlexLLMGen DeepSpeed Reduction Utils
- Implementation:Explodinggradients Ragas DSPyOptimizer Class
- Implementation:Volcengine Verl Compute GRPO Outcome Advantage
Heuristics
- Heuristic:BerriAI Litellm Retry Backoff Strategy
- Heuristic:FlagOpen FlagEmbedding Length Sorted Batching
- Heuristic:Langgenius Dify Token Refresh Loop Prevention
- Heuristic:Apache Airflow Scheduler Performance Tuning
- Heuristic:Microsoft LoRA Selective LoRA QV Only
- Heuristic:Teamcapybara Capybara Async Waiting And Retry
- Heuristic:Nightwatchjs Nightwatch Timeout And Retry Tuning
- Heuristic:Vespa engine Vespa Warning Deprecated Cloud API Constructors
- Heuristic:OpenBMB UltraFeedback Principle Distribution Tuning
- Heuristic:Marker Inc Korea AutoRAG OpenAI Rate Limit Mitigation
Environments
- Environment:Intel Ipex llm RAG LangChain Environment
- Environment:Huggingface Datasets Lance Dependencies
- Environment:Marker Inc Korea AutoRAG Japanese NLP Dependencies
- Environment:NVIDIA DALI TensorFlow Environment
- Environment:Datajuicer Data juicer Ray Cluster Environment
- Environment:ArroyoSystems Arroyo Kubernetes Deployment
- Environment:Openai Openai agents python MCP Dependencies
- Environment:OpenHands OpenHands SaaS Server Environment
- Environment:Promptfoo Promptfoo Node Runtime
- Environment:Mlfoundations Open flamingo WebDataset Training Dependencies