Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — full-stack AI/ML coding agent
- Leeroopedia MCP — knowledge search for any coding agent
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Microsoft Autogen Graph Based Agent Orchestration
- Workflow:Alibaba ROLL DPO Training Pipeline
- Workflow:Neuml Txtai Workflow Orchestration
- Workflow:Datahub project Datahub Metadata Actions Pipeline
- Workflow:Scikit learn Scikit learn Cross Validation Evaluation
- Workflow:Roboflow Rf detr Object Detection Inference
- Workflow:ArroyoSystems Arroyo Checkpoint Recovery
- Workflow:MaterializeInc Materialize Release Process
- Workflow:Alibaba MNN Model Conversion Pipeline
- Workflow:Nautechsystems Nautilus trader Backtest with BacktestEngine
Principles
- Principle:Intel Ipex llm DPO Model Loading
- Principle:Ucbepic Docetl Dataset Upload And Operation Creation
- Principle:Unslothai Unsloth GRPO Reinforcement Learning
- Principle:Pytorch Serve Multimodal Inference
- Principle:Openclaw Openclaw Message Ingestion
- Principle:Google deepmind Mujoco Vectorized Simulation
- Principle:Rapidsai Cuml Cluster Model Fitting
- Principle:TobikoData Sqlmesh Github Actions Integration
- Principle:Promptfoo Promptfoo External Data Source Integration
- Principle:Explodinggradients Ragas DSPy Prompt Optimization
Implementations
- Implementation:Kserve Kserve InferenceGraph Full CRD
- Implementation:Ucbepic Docetl NaturalLanguagePipelineDialog
- Implementation:DevExpress Testcafe CompileClientFunction
- Implementation:NVIDIA TransformerEngine JAX Sharding
- Implementation:Speechbrain Speechbrain Train L2I
- Implementation:Getgauge Taiko EmulateNetwork
- Implementation:ArroyoSystems Arroyo Kafka Source
- Implementation:Neuml Txtai PGText Scoring
- Implementation:FMInference FlexLLMGen Bench HF Suite
- Implementation:Vllm project Vllm EngineArgs LoRA Config
Heuristics
- Heuristic:OpenGVLab InternVL LoRA Alpha Scaling
- Heuristic:Alibaba ROLL Sequence Packing Alignment
- Heuristic:Fastai Fastbook Discriminative Learning Rates
- Heuristic:NVIDIA DALI Thread Affinity Optimization
- Heuristic:Farama Foundation Gymnasium Render Mode Selection Guide
- Heuristic:DataExpert io Data engineer handbook Watermark Late Arrival Tolerance
- Heuristic:Alibaba ROLL PPO Clipping Defaults
- Heuristic:Zai org CogVideo DeepSpeed Checkpoint Conversion Tips
- Heuristic:Mlc ai Mlc llm Metal KV Cache Capacity Limit
- Heuristic:Alibaba MNN NC4HW4 Data Layout
Environments
- Environment:Apache Kafka Committer Tools Environment
- Environment:Snorkel team Snorkel PyTorch
- Environment:Apache Dolphinscheduler Netty Runtime
- Environment:Cohere ai Cohere python Cohere API Credentials
- Environment:Mlfoundations Open flamingo WebDataset Training Dependencies
- Environment:Huggingface Transformers Flash Attention 2 Env
- Environment:DevExpress Testcafe Firefox Marionette
- Environment:Getgauge Taiko Chromium Browser
- Environment:Treeverse LakeFS Go Runtime Environment
- Environment:OWASP Www project top 10 for large language model applications Pre Commit Hooks Environment