Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Bentoml BentoML BentoCloud Deployment
- Workflow:PrefectHQ Prefect AI Data Analyst Agent
- Workflow:Apache Flink Async Sink Lifecycle
- Workflow:Openai Whisper Word Level Timestamps
- Workflow:Ggml org Llama cpp OpenAI Compatible Server
- Workflow:Langchain ai Langgraph Human in the Loop Agent
- Workflow:Nightwatchjs Nightwatch Component Testing
- Workflow:Spotify Luigi Database Ingestion Pipeline
- Workflow:Lucidrains X transformers Encoder Decoder Sequence to Sequence
- Workflow:Huggingface Peft DreamBooth LoRA Diffusion
Principles
- Principle:CrewAIInc CrewAI Knowledge Ingestion
- Principle:Spotify Luigi Web Visualiser Monitoring
- Principle:Langchain ai Langgraph Cache Backend Selection
- Principle:Cohere ai Cohere python Streaming Chat Request
- Principle:Speechbrain Speechbrain Speaker Verification Scoring
- Principle:Alibaba MNN Weight Quantization
- Principle:Princeton nlp Tree of thought llm BFS Tree Search
- Principle:Interpretml Interpret EBM JSON Serialization
- Principle:SeldonIO Seldon core Usage Metrics Publishing
- Principle:Dagster io Dagster Project Scaffolding
Implementations
- Implementation:InternLM Lmdeploy RmsNorm
- Implementation:EvolvingLMMs Lab Lmms eval IFEval Utils
- Implementation:Bitsandbytes foundation Bitsandbytes PagedAdamW8bit
- Implementation:Ollama Ollama Llama Sampling API
- Implementation:Huggingface Transformers Grafana Benchmark Dashboard
- Implementation:Infiniflow Ragflow Time Utils
- Implementation:Huggingface Datatrove BaseExtractor
- Implementation:Lance format Lance VectorThroughputBench
- Implementation:Datajuicer Data juicer LanguageIDScoreFilter
- Implementation:Pola rs Polars Buffer
Heuristics
- Heuristic:Pytorch Serve CPU Performance Tuning
- Heuristic:Mlflow Mlflow Async Logging Best Practices
- Heuristic:Datahub project Datahub Gradle Formatting Over Direct Tools
- Heuristic:Apache Paimon File Sizing and Split Planning
- Heuristic:Bigscience workshop Petals Randomized Rebalancing Intervals
- Heuristic:Ollama Ollama Sampling Numerical Stability
- Heuristic:Huggingface Datatrove VLLM Startup Optimization
- Heuristic:Marker Inc Korea AutoRAG Hybrid Retrieval Score Normalization
- Heuristic:Obss Sahi Overlap Ratio Selection
- Heuristic:FMInference FlexLLMGen Pin Memory Tradeoffs
Environments
- Environment:Nautechsystems Nautilus trader Binance API Credentials
- Environment:Interpretml Interpret Native Libebm Environment
- Environment:Anthropics Anthropic sdk python Azure Foundry Environment
- Environment:LMCache LMCache VLLM Serving Engine
- Environment:NVIDIA DALI PyTorch Environment
- Environment:Huggingface Peft Python Core Dependencies
- Environment:Run llama Llama index OpenAI API Configuration
- Environment:Infiniflow Ragflow GPU CUDA Environment
- Environment:SeldonIO Seldon core Python ML Dependencies Environment
- Environment:Google deepmind Dm control EGL Headless Rendering