Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Kubeflow Kubeflow Contributing To Kubeflow
- Workflow:Scikit learn Scikit learn Data Preprocessing Pipeline
- Workflow:Spcl Graph of thoughts GoT Sorting Pipeline
- Workflow:EvolvingLMMs Lab Lmms eval Distributed Multi GPU Evaluation
- Workflow:Scikit learn Scikit learn Cross Validation Evaluation
- Workflow:Farama Foundation Gymnasium Agent Evaluation Recording
- Workflow:Open compass VLMEvalKit Adding Custom Benchmark
- Workflow:Lance format Lance Table Optimization
- Workflow:Vllm project Vllm Speculative Decoding
- Workflow:Predibase Lorax Single LoRA Inference
Principles
- Principle:Sdv dev SDV Cardinality Analysis
- Principle:Apache Druid Streaming Schema Spec
- Principle:ClickHouse ClickHouse ELF Program Header Caching
- Principle:Apache Spark Source Compilation
- Principle:CrewAIInc CrewAI Training Execution
- Principle:FlowiseAI Flowise AI Flow Generation
- Principle:Datahub project Datahub Deployment Verification
- Principle:Apache Beam Parallel Execution
- Principle:Ggml org Llama cpp Server Configuration
- Principle:Unslothai Unsloth AIME Evaluation
Implementations
- Implementation:Spotify Luigi Marker Table Check
- Implementation:Groq Groq python Transcription Response
- Implementation:Online ml River Compat RiverToSklearn
- Implementation:Ggml org Ggml Cann backend api
- Implementation:Neuml Txtai Annoy ANN
- Implementation:Intel Ipex llm Pipeline Parallel Serving
- Implementation:Guardrails ai Guardrails Schema Validator
- Implementation:OpenHands OpenHands GithubIssue Initialize Conversation
- Implementation:Vespa engine Vespa Searchlib ABI Spec
- Implementation:Triton inference server Server L0 Sequence Batcher Test
Heuristics
- Heuristic:ChenghaoMou Text dedup Fingerprint Batch Size One
- Heuristic:Neuml Txtai Model Quantization Defaults
- Heuristic:Princeton nlp SimPO Multi Seed Diversity
- Heuristic:Duckdb Duckdb Version Sync Across Files
- Heuristic:Spotify Luigi Atomic File Writes
- Heuristic:OpenGVLab InternVL Packed Training Buffer Management
- Heuristic:Microsoft LoRA LoRA Init Strategy
- Heuristic:OpenGVLab InternVL Pixel Shuffle Downsampling
- Heuristic:Fastai Fastbook Progressive Resizing
- Heuristic:Open compass VLMEvalKit SKIP ERR For Graceful Inference
Environments
- Environment:Marker Inc Korea AutoRAG API Keys And Credentials
- Environment:Vespa engine Vespa Docker OCI Container Runtime
- Environment:Neuml Txtai GPU Accelerator Detection
- Environment:Sgl project Sglang CUDA Runtime
- Environment:Protectai Modelscan Python Core Runtime
- Environment:Allenai Open instruct Beaker Cluster
- Environment:Unstructured IO Unstructured OpenAI API
- Environment:OWASP Www project top 10 for large language model applications GenAI Red Team Environment
- Environment:Testtimescaling Testtimescaling github io GitHub Actions Runner
- Environment:LMCache LMCache Python Runtime