Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Microsoft DeepSpeedExamples ZeRO Inference
- Workflow:Nightwatchjs Nightwatch Custom Commands And Assertions
- Workflow:Kornia Kornia Differentiable Image Augmentation
- Workflow:Astronomer Astronomer cosmos Local dbt DAG rendering
- Workflow:Apache Druid Batch Data Ingestion
- Workflow:Mage ai Mage ai Building a New Source Connector
- Workflow:Bigscience workshop Petals Prompt Tuning Chatbot
- Workflow:DataTalksClub Data engineering zoomcamp Docker PostgreSQL Data Ingestion
- Workflow:Volcengine Verl Multi Turn Tool Use Training
- Workflow:Unslothai Unsloth Vision Model Finetuning
Principles
- Principle:Arize ai Phoenix Evaluation Data Preparation
- Principle:LaurentMazare Tch rs Backpropagation Training
- Principle:EvolvingLMMs Lab Lmms eval Task Utility Functions
- Principle:Huggingface Transformers Training Configuration
- Principle:ClickHouse ClickHouse Hex Binary Encoding
- Principle:Huggingface Datatrove Line Level Statistics
- Principle:CrewAIInc CrewAI Tool Implementation
- Principle:OpenRLHF OpenRLHF Rejection Sampling
- Principle:Deepseek ai Janus Image Loading and Preprocessing
- Principle:OpenHands OpenHands Frontend Serving
Implementations
- Implementation:SeleniumHQ Selenium Closure Dom Selection
- Implementation:DataExpert io Data engineer handbook Tumble Over Window
- Implementation:OpenHands OpenHands SaasSettingsStore
- Implementation:DataTalksClub Data engineering zoomcamp Airflow Homework Solution
- Implementation:CrewAIInc CrewAI Invoke Automation Tool
- Implementation:FMInference FlexLLMGen OptLM Init
- Implementation:Recommenders team Recommenders Python Chrono Split
- Implementation:Evidentlyai Evidently Legacy Contains Link Feature
- Implementation:Facebookresearch Habitat lab PointNavResNetPolicy from config
- Implementation:Microsoft Playwright Server Input
Heuristics
- Heuristic:Run llama Llama index Worker Count Configuration
- Heuristic:Farama Foundation Gymnasium Action Space Normalization Tip
- Heuristic:Deepspeedai DeepSpeed FP16 Convergence Tips
- Heuristic:Huggingface Peft LoRA Default Configuration
- Heuristic:Allenai Open instruct GPU Memory Utilization
- Heuristic:Eric mitchell Direct preference optimization Activation Checkpointing Memory
- Heuristic:Google deepmind Mujoco Mesh Quality For Collision
- Heuristic:OWASP Www project top 10 for large language model applications Sandbox Containerization Pattern
- Heuristic:Risingwavelabs Risingwave LSM Compaction Tuning
- Heuristic:Hiyouga LLaMA Factory LoRA DDP Configuration
Environments
- Environment:Apache Dolphinscheduler Netty Runtime
- Environment:Deepspeedai DeepSpeed XPU Environment
- Environment:NVIDIA DALI TensorFlow Environment
- Environment:Ggml org Llama cpp CUDA GPU Environment
- Environment:BerriAI Litellm Redis Cache Backend
- Environment:Sdv dev SDV Python Runtime
- Environment:Arize ai Phoenix LLM Provider SDKs
- Environment:Lm sys FastChat LoRA QLoRA Training Environment
- Environment:Snorkel team Snorkel SpaCy NLP
- Environment:Huggingface Alignment handbook BitsAndBytes CUDA