Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Cypress io Cypress CI Pipeline Integration
- Workflow:Unstructured IO Unstructured Document Partitioning
- Workflow:SeldonIO Seldon core Inference Pipeline
- Workflow:Sgl project Sglang Offline Batch Inference
- Workflow:Princeton nlp Tree of thought llm Adding new task
- Workflow:Cohere ai Cohere python Model Finetuning
- Workflow:Microsoft BIPIA White Box Defense Finetuning
- Workflow:Predibase Lorax Single LoRA Inference
- Workflow:Ggml org Llama cpp Multimodal Inference
- Workflow:Mlc ai Web llm Text Embeddings And RAG
Principles
- Principle:Huggingface Datasets Dataset Metadata Configuration
- Principle:Deepspeedai DeepSpeed Op Builder System
- Principle:FMInference FlexLLMGen Distributed Pipeline Initialization
- Principle:OpenGVLab InternVL Distributed Evaluation
- Principle:Run llama Llama index Batch Evaluation Execution
- Principle:Langchain ai Langgraph Server Runtime Context
- Principle:Recommenders team Recommenders SAR Algorithm
- Principle:Apache Kafka Topic Deletion
- Principle:Neuml Txtai Agent Configuration
- Principle:Sgl project Sglang Sampling Parameters Preparation
Implementations
- Implementation:Lucidrains X transformers XTransformer Init
- Implementation:Openai Openai python Eval List Response
- Implementation:Wandb Weave Publish PyPI Release
- Implementation:Online ml River Metrics AdjustedRand
- Implementation:Scikit learn Scikit learn ArffParser
- Implementation:Apache Druid SchemaColumnList
- Implementation:Risingwavelabs Risingwave MetaClient
- Implementation:Apache Paimon MathUtils
- Implementation:Google deepmind Mujoco JAX Experimental FFI
- Implementation:Open compass VLMEvalKit MMHelix WordLadder Eval
Heuristics
- Heuristic:AUTOMATIC1111 Stable diffusion webui NaN Detection And Precision Fixes
- Heuristic:Langchain ai Langchain Warning Deprecated AnthropicLLM
- Heuristic:Apache Shardingsphere Worker ID Reservation Strategy
- Heuristic:LLMBook zh LLMBook zh github io IGNORE INDEX Loss Masking
- Heuristic:FMInference FlexLLMGen Offloading Percent Tuning
- Heuristic:Arize ai Phoenix Notebook Event Loop Patching
- Heuristic:Openai Openai node Retry Backoff Configuration
- Heuristic:NVIDIA DALI Batch Size Tuning
- Heuristic:TA Lib Ta lib python Compatibility Mode Switching
- Heuristic:Huggingface Datasets Batch Size Optimization
Environments
- Environment:Alibaba ROLL Megatron Training Environment
- Environment:Risingwavelabs Risingwave Dashboard Node Environment
- Environment:Huggingface Transformers Python 310 Runtime
- Environment:Unstructured IO Unstructured OpenAI API
- Environment:Vllm project Vllm CUDA
- Environment:Apache Airflow Python Runtime Environment
- Environment:Bentoml BentoML Triton Inference Server
- Environment:Hpcaitech ColossalAI CUDA GPU Environment
- Environment:Volcengine Verl Ray Distributed Environment
- Environment:Deepspeedai DeepSpeed Multi Accelerator Environment