Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Ucbepic Docetl Long Document Chunking
- Workflow:Apache Airflow Kubernetes Deployment via Helm
- Workflow:Openai Openai python Realtime Conversation
- Workflow:ClickHouse ClickHouse Running Stateless Tests
- Workflow:Microsoft Agent framework Multi Agent Sequential Orchestration
- Workflow:Hiyouga LLaMA Factory Full Parameter SFT
- Workflow:Evidentlyai Evidently ML Model Quality Report
- Workflow:Vibrantlabsai Ragas Experiment Driven Development
- Workflow:Microsoft Onnxruntime Nodejs Inference
- Workflow:Neuml Txtai Semantic Search Pipeline
Principles
- Principle:Duckdb Duckdb Declarative Benchmark Execution
- Principle:CARLA simulator Carla NPC Vehicle Configuration
- Principle:Microsoft Autogen Agent Specialization
- Principle:Tensorflow Serving Version Policy Configuration
- Principle:Webdriverio Webdriverio Selector Processing
- Principle:Microsoft DeepSpeedExamples Distributed Checkpoint Saving
- Principle:ClickHouse ClickHouse Data Lake Testing
- Principle:Sdv dev SDV Cardinality Analysis
- Principle:Fastai Fastbook Language Model Data
- Principle:Unslothai Unsloth Vision Model Loading
Implementations
- Implementation:SeleniumHQ Selenium Platform
- Implementation:FlagOpen FlagEmbedding Reinforced IR Model
- Implementation:Apache Druid SegmentsView
- Implementation:Huggingface Transformers Setup Py
- Implementation:ClickHouse ClickHouse Poco RemoteSyslogListener
- Implementation:Open compass VLMEvalKit VGRPBench Aquarium
- Implementation:Huggingface Datatrove DocumentTokenizerMerger
- Implementation:Lance format Lance LegacyBitpackEncoding
- Implementation:Huggingface Datasets HDF5 Builder
- Implementation:Ggml org Ggml Cpu kleidiai kernels
Heuristics
- Heuristic:Haotian liu LLaVA Image Aspect Ratio Padding Strategy
- Heuristic:Liu00222 Open Prompt Injection PPL Threshold Tuning
- Heuristic:Microsoft Playwright Test Stability Practices
- Heuristic:Apache Shardingsphere Worker ID Reservation Strategy
- Heuristic:Eric mitchell Direct preference optimization TF32 Matmul Precision
- Heuristic:Lucidrains X transformers Rotary Position Embedding Selection
- Heuristic:Cleanlab Cleanlab Multiprocessing Platform Strategy
- Heuristic:Alibaba ROLL Reward Clipping Normalization
- Heuristic:Shiyu coder Kronos Two Stage Finetuning Strategy
- Heuristic:Kubeflow Pipelines Cache Staleness In Recursive Pipelines
Environments
- Environment:Apache Paimon Cloud Storage Credentials
- Environment:Vllm project Vllm Intel XPU
- Environment:Langchain ai Langgraph Docker Deployment Environment
- Environment:BerriAI Litellm Docker Deployment
- Environment:Huggingface Datatrove IO Dependencies
- Environment:LMCache LMCache NIXL Transfer Library
- Environment:Wandb Weave Trace Server Infrastructure
- Environment:OpenRLHF OpenRLHF Ray Distributed Environment
- Environment:Ggml org Llama cpp Python Conversion Environment
- Environment:Huggingface Diffusers Training Environment