Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Apache Druid Batch Data Ingestion
- Workflow:Rapidsai Cuml Dimensionality Reduction
- Workflow:LMCache LMCache Disaggregated Prefill
- Workflow:Pytorch Serve Model Deployment
- Workflow:Ollama Ollama Safetensors To GGUF Conversion
- Workflow:Pola rs Polars Streaming Large Dataset Processing
- Workflow:OpenRLHF OpenRLHF SFT Training
- Workflow:Huggingface Trl Direct Preference Optimization
- Workflow:Risingwavelabs Risingwave Docker Deployment
- Workflow:Interpretml Interpret Blackbox Model Explanation
Principles
- Principle:Eventual Inc Daft Delta Lake Writing
- Principle:Eventual Inc Daft Iceberg Writing
- Principle:Tencent Ncnn Vulkan Inference Configuration
- Principle:EvolvingLMMs Lab Lmms eval Task Selection
- Principle:PrefectHQ Prefect Webhook Driven Flow Resumption
- Principle:Mlflow Mlflow Evaluation Dataset Preparation
- Principle:Ollama Ollama Compute Graph System
- Principle:Promptfoo Promptfoo Quality Gate Enforcement
- Principle:Ollama Ollama MLXRunner Architecture
- Principle:Cleanlab Cleanlab Label Issue Ordering
Implementations
- Implementation:Facebookresearch Habitat lab InfoDict Utils
- Implementation:Openai Openai python Vector Store Create Params
- Implementation:Hiyouga LLaMA Factory KTransformers Integration
- Implementation:FlagOpen FlagEmbedding Matryoshka Self Distillation Modeling
- Implementation:Openai Openai python File Object Model
- Implementation:Langchain ai Langgraph Public Constants
- Implementation:Cohere ai Cohere python AwsClassification
- Implementation:InternLM Lmdeploy RequestLogger
- Implementation:ThreeSR Awesome Inference Time Scaling Get Paper Info Function
- Implementation:Openai Openai python Lazy Proxy
Heuristics
- Heuristic:Alibaba ROLL Reward Clipping Normalization
- Heuristic:Microsoft LoRA LoRA Init Strategy
- Heuristic:Danijar Dreamerv3 Free Nats KL Thresholding
- Heuristic:Mlfoundations Open flamingo Gradient Clipping Max Norm
- Heuristic:Mlfoundations Open flamingo Deterministic Shard Shuffling
- Heuristic:Sail sg LongSpec Tree Shape Configuration
- Heuristic:LMCache LMCache Chunk Size And Default Config
- Heuristic:Mlc ai Web llm Grammar Matcher Reuse
- Heuristic:ClickHouse ClickHouse Jemalloc Production Requirement
- Heuristic:Unstructured IO Unstructured PDF Element Sorting
Environments
- Environment:Testtimescaling Testtimescaling github io GitHub Actions Runner
- Environment:Microsoft BIPIA OpenAI API Environment
- Environment:Mbzuai oryx Awesome LLM Post training Git CLI
- Environment:OWASP Www project top 10 for large language model applications PR Description Generator Runtime
- Environment:Helicone Helicone Python ClickHouse Migrations
- Environment:Vibrantlabsai Ragas Optional NLP Metrics Environment
- Environment:SeldonIO Seldon core Kafka Messaging Environment
- Environment:Deepspeedai DeepSpeed CUDA GPU Environment
- Environment:Alibaba ROLL SGLang Inference Environment
- Environment:Gretelai Gretel synthetics PyTorch CUDA Environment