Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Explodinggradients Ragas LLM Benchmarking
- Workflow:Risingwavelabs Risingwave Iceberg Lakehouse Ingestion
- Workflow:Eric mitchell Direct preference optimization SFT Training
- Workflow:Tensorflow Tfjs GPT2 Text Generation
- Workflow:Scikit learn Scikit learn Cross Validation Evaluation
- Workflow:AnswerDotAI RAGatouille ColBERT Training
- Workflow:Deepset ai Haystack Extractive QA Pipeline
- Workflow:Apache Spark Application Submission
- Workflow:Microsoft Autogen Studio Team Deployment
- Workflow:NVIDIA NeMo Aligner RLHF PPO Training
Principles
- Principle:Huggingface Peft Quantized Model Preparation
- Principle:ClickHouse ClickHouse Glibc Symbol Replacement
- Principle:Unslothai Unsloth Reward Function Design
- Principle:Predibase Lorax Stateless Conversation Management
- Principle:Confident ai Deepeval Exception Hierarchy Design
- Principle:Openai Openai agents python Documentation Build Pipeline
- Principle:AUTOMATIC1111 Stable diffusion webui Configuration Management
- Principle:Mlfoundations Open flamingo FSDP Model Wrapping
- Principle:Apache Airflow Security Review Planning
- Principle:Datahub project Datahub Pipeline Execution
Implementations
- Implementation:Speechbrain Speechbrain Prepare Switchboard Transformer
- Implementation:InternLM Lmdeploy AnomalyHandler
- Implementation:Anthropics Anthropic sdk python Exceptions
- Implementation:Duckdb Duckdb Mbedtls ECP
- Implementation:ARISE Initiative Robosuite Observables
- Implementation:Openai Openai python Completion Create Params
- Implementation:Scikit learn Scikit learn LinearModelModule
- Implementation:FlowiseAI Flowise EvaluationResultSideDrawer
- Implementation:Huggingface Datasets ImageFolder Builder
- Implementation:NVIDIA DALI C API V2 Ref Counting
Heuristics
- Heuristic:AnswerDotAI RAGatouille FAISS Vs PyTorch KMeans Indexing
- Heuristic:Hiyouga LLaMA Factory Mixed Precision Training Tips
- Heuristic:Pola rs Polars Collect All For Diverging Queries
- Heuristic:Datahub project Datahub Emitter Selection Strategy
- Heuristic:Kubeflow Pipelines Component URL Commit SHA Pinning
- Heuristic:Apache Shardingsphere Version Cleanup After Switch
- Heuristic:Arize ai Phoenix Notebook Event Loop Patching
- Heuristic:Vespa engine Vespa Warning Deprecated Cloud API Constructors
- Heuristic:PacktPublishing LLM Engineers Handbook Temperature Selection By Task
- Heuristic:OpenRLHF OpenRLHF vLLM Embedding Resize Warning
Environments
- Environment:Fastai Fastbook NLP SpaCy Environment
- Environment:Huggingface Alignment handbook PyTorch CUDA
- Environment:MaterializeInc Materialize Kubernetes Helm Runtime
- Environment:Vespa engine Vespa Docker OCI Container Runtime
- Environment:Farama Foundation Gymnasium Video Recording Dependencies
- Environment:Wandb Weave LLM Integration Dependencies
- Environment:Allenai Open instruct CUDA GPU Training
- Environment:Huggingface Datatrove S3 Storage Environment
- Environment:Mit han lab Llm awq Flash Attention Environment
- Environment:Openai Whisper Triton