Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Google research Deduplicate text datasets Wiki40B TFDS deduplication
- Workflow:Ucbepic Docetl YAML Pipeline Execution
- Workflow:CARLA simulator Carla Simulation Setup and First Steps
- Workflow:Kubeflow Pipelines XGBoost Training Pipeline
- Workflow:Run llama Llama index ReAct Agent
- Workflow:Helicone Helicone LLM Response Normalization
- Workflow:Sgl project Sglang Offline Batch Inference
- Workflow:Nightwatchjs Nightwatch Cucumber BDD Integration
- Workflow:VainF Torch Pruning Object Detection Pruning
- Workflow:Princeton nlp Tree of thought llm ToT BFS experiment
Principles
- Principle:Togethercomputer Together python Image Prompt Construction
- Principle:Huggingface Datatrove MIME Type Filtering
- Principle:Deepspeedai DeepSpeed RLHF Checkpointing
- Principle:Online ml River Optimizer Configuration
- Principle:Apache Hudi Read Environment Configuration
- Principle:Langfuse Langfuse Ingestion Queue Dispatch
- Principle:ARISE Initiative Robomimic Observation Initialization
- Principle:Webdriverio Webdriverio Framework Abstraction
- Principle:Neuml Txtai Semantic Search
- Principle:Datajuicer Data juicer Data Grouping
Implementations
- Implementation:Triton inference server Server L0 Batcher Test
- Implementation:NVIDIA TransformerEngine Float8 Storage
- Implementation:Triton inference server Server L0 Socket Test
- Implementation:Allenai Open instruct Reward Modeling Main
- Implementation:Iterative Dvc Api Scm
- Implementation:Microsoft Autogen Studio MCP Sidebar
- Implementation:Datajuicer Data juicer Init Configs
- Implementation:AUTOMATIC1111 Stable diffusion webui Extra Options Section
- Implementation:Intel Ipex llm Transformers Trainer LoRA
- Implementation:MarketSquare Robotframework browser Rfbrowser Init
Heuristics
- Heuristic:Huggingface Datatrove VLLM Startup Optimization
- Heuristic:ArroyoSystems Arroyo Stateful Operator TTL
- Heuristic:Alibaba MNN Memory Mode Selection
- Heuristic:SeleniumHQ Selenium Bazel Hermetic Build Requirement
- Heuristic:Microsoft BIPIA LLAMA Pad Token Workaround
- Heuristic:Lm sys FastChat Flash Attention GPU Requirements
- Heuristic:Openclaw Openclaw Retry With Exponential Backoff
- Heuristic:Google research Deduplicate text datasets Variable Width Pointer Optimization
- Heuristic:Mlfoundations Open flamingo Gradient Clipping Max Norm
- Heuristic:Unstructured IO Unstructured Chunk Size Tuning
Environments
- Environment:Marker Inc Korea AutoRAG VLLM Environment
- Environment:Evidentlyai Evidently Spark Engine Environment
- Environment:Kserve Kserve Knative Serving
- Environment:ARISE Initiative Robomimic HDF5 Data Dependencies
- Environment:Eventual Inc Daft Python PyArrow Core
- Environment:Vespa engine Vespa FNET Transport Config
- Environment:Langfuse Langfuse ClickHouse Analytics
- Environment:Pyro ppl Pyro CUDA GPU Acceleration
- Environment:Mlc ai Web llm Chrome Extension Manifest V3
- Environment:InternLM Lmdeploy CUDA GPU Runtime