Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Apache Hudi Flink Batch Incremental Read
- Workflow:Promptfoo Promptfoo CI CD Integration
- Workflow:DevExpress Testcafe CLI Test Execution
- Workflow:Huggingface Datasets Hub Publishing
- Workflow:Mage ai Mage ai SQL Database Source Extraction
- Workflow:Neuml Txtai RAG Pipeline
- Workflow:Datahub project Datahub Spark Lineage Capture
- Workflow:Anthropics Anthropic sdk python Extended Thinking Reasoning
- Workflow:Anthropics Anthropic sdk python Structured Output Extraction
- Workflow:LaurentMazare Tch rs MNIST Training
Principles
- Principle:Volcengine Verl Environment Setup
- Principle:Protectai Modelscan Middleware Pipeline
- Principle:AUTOMATIC1111 Stable diffusion webui Embedding creation
- Principle:Teamcapybara Capybara Selector Modification
- Principle:Togethercomputer Together python Embedding Generation
- Principle:Iterative Dvc Content Hashing
- Principle:Rapidsai Cuml Incremental Learning
- Principle:Lm sys FastChat Train Test Data Splitting
- Principle:DataExpert io Data engineer handbook Tumbling Window
- Principle:Marker Inc Korea AutoRAG Trial Summary And Dashboard
Implementations
- Implementation:Kserve Kserve RDMA Network Configuration
- Implementation:Langfuse Langfuse API Metrics Schema
- Implementation:Speechbrain Speechbrain Create Aishell1Mix Metadata
- Implementation:DevExpress Testcafe BrowserClient CDP
- Implementation:Volcengine Verl Multimodal Rollout Request
- Implementation:Fede1024 Rust rdkafka MockCluster Create Topic
- Implementation:Openai Openai python Response Image Gen Call Completed
- Implementation:CrewAIInc CrewAI Tool Assignment Config
- Implementation:Alibaba MNN Protobuf Generated Table Driven Lite H
- Implementation:Bentoml BentoML Project Configuration
Heuristics
- Heuristic:Hiyouga LLaMA Factory LoRA DDP Configuration
- Heuristic:Google research Deduplicate text datasets HACKSIZE Overlap Buffer
- Heuristic:Microsoft Playwright Timeout Configuration Tips
- Heuristic:Unslothai Unsloth LoRA Rank Selection
- Heuristic:Microsoft Autogen Warning Deprecated JSON Env Files
- Heuristic:Isaac sim IsaacGymEnvs JIT Profiling Optimization
- Heuristic:Helicone Helicone Anthropic Cache Double Count Prevention
- Heuristic:Duckdb Duckdb Sanitizer Configuration
- Heuristic:Vibrantlabsai Ragas Concurrency And Retry Configuration
- Heuristic:Online ml River HST Feature Scaling Requirement
Environments
- Environment:Ggml org Ggml Metal GPU Environment
- Environment:Tencent Ncnn PyTorch Environment
- Environment:PacktPublishing LLM Engineers Handbook Docker MongoDB Qdrant Infrastructure
- Environment:Mlfoundations Open flamingo PyTorch CUDA Distributed
- Environment:Cleanlab Cleanlab Datalab Dependencies
- Environment:Kserve Kserve GPU Accelerator
- Environment:Datajuicer Data juicer Ray Cluster Environment
- Environment:Allenai Open instruct vLLM Inference
- Environment:Huggingface Trl vLLM Generation Environment
- Environment:Huggingface Alignment handbook Python PEFT