Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Huggingface Open r1 SFT Distillation
- Workflow:Apache Druid Batch Data Ingestion
- Workflow:Langgenius Dify Application Creation
- Workflow:Iamhankai Forest of Thought CGDM Post Processing
- Workflow:Nightwatchjs Nightwatch Custom Commands And Assertions
- Workflow:Lm sys FastChat Distributed Model Serving
- Workflow:BerriAI Litellm Proxy Server Deployment
- Workflow:Langfuse Langfuse Otel ingestion pipeline
- Workflow:OpenRLHF OpenRLHF Rejection Sampling
- Workflow:LMCache LMCache KV Cache Offloading
Principles
- Principle:Sktime Pytorch forecasting Learning Rate Finding
- Principle:Ggml org Ggml BPE Tokenization
- Principle:ClickHouse ClickHouse Code Review Process
- Principle:Sktime Pytorch forecasting V2 Data Pipeline
- Principle:Spotify Luigi Spark Job Definition
- Principle:Interpretml Interpret Feature Binning And Discretization
- Principle:HKUDS AI Trader LLM Invocation
- Principle:Vllm project Vllm Streaming Response Handling
- Principle:Heibaiying BigData Notes Spark Session Creation
- Principle:Apache Spark Test Orchestration
Implementations
- Implementation:OpenGVLab InternVL LLaVA LLaMA Model
- Implementation:Recommenders team Recommenders EmbDotBias Score
- Implementation:Ggml org Llama cpp Preset
- Implementation:Facebookresearch Habitat lab GuiInput
- Implementation:NVIDIA DALI ImageNet Synsets
- Implementation:Openai Openai python Image Gen Completed Event
- Implementation:Tensorflow Tfjs Python Inference
- Implementation:Openai Evals Pip Install Evals
- Implementation:Open compass VLMEvalKit SArena FID
- Implementation:Alibaba MNN ExecutorScope
Heuristics
- Heuristic:Nightwatchjs Nightwatch Safari Parallel Limitation
- Heuristic:EvolvingLMMs Lab Lmms eval Limit Flag Testing Only
- Heuristic:Treeverse LakeFS Retry Backoff Configuration
- Heuristic:Huggingface Datasets Batch Size Optimization
- Heuristic:Lm sys FastChat Flash Attention GPU Requirements
- Heuristic:Deepset ai Haystack Document Splitting Defaults
- Heuristic:Ollama Ollama Download Retry Strategy
- Heuristic:Ollama Ollama Quantization Layer Selection
- Heuristic:ArroyoSystems Arroyo Parallelism Configuration
- Heuristic:VainF Torch Pruning Over Pruning Prevention
Environments
- Environment:Huggingface Open r1 vLLM Server
- Environment:Tensorflow Serving Python Client Environment
- Environment:Sail sg LongSpec Training Environment
- Environment:TobikoData Sqlmesh Snowflake Connection
- Environment:Ucbepic Docetl Frontend Node Environment
- Environment:Marker Inc Korea AutoRAG Korean NLP Dependencies
- Environment:Promptfoo Promptfoo SQLite Database
- Environment:Explodinggradients Ragas Google Drive Backend Environment
- Environment:Turboderp org Exllamav2 CUDA GPU Runtime
- Environment:Microsoft BIPIA DeepSpeed Finetuning Environment