Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Huggingface Open r1 GRPO Reasoning Training
- Workflow:Vibrantlabsai Ragas RAG Evaluation
- Workflow:Roboflow Rf detr Roboflow Deployment
- Workflow:Turboderp org Exllamav2 LoRA Adapter Inference
- Workflow:FlagOpen FlagEmbedding Reranker Inference
- Workflow:Kornia Kornia ONNX Model Pipeline
- Workflow:NVIDIA DALI Custom Operator Development
- Workflow:Openai Evals Creating a model graded eval
- Workflow:Huggingface Diffusers Checkpoint Conversion
- Workflow:Datahub project Datahub Docker Quickstart Deployment
Principles
- Principle:Microsoft Autogen Graph Flow Configuration
- Principle:Getgauge Taiko Gauge Step Implementation
- Principle:Openai Openai python Image Generation
- Principle:Kserve Kserve Authentication Integration
- Principle:Scikit learn Scikit learn Online Learning
- Principle:NVIDIA NeMo Aligner PPO Actor Critic Setup
- Principle:Webdriverio Webdriverio Plugin Resolution
- Principle:Langfuse Langfuse Prompt Compilation with Variables
- Principle:Microsoft Agent framework Workflow Event Streaming
- Principle:Langgenius Dify Explore Feature
Implementations
- Implementation:Lance format Lance Java DatasetDeltaBuilder
- Implementation:Tencent Ncnn SCRFD CrowdHuman Example
- Implementation:Huggingface Datasets Tqdm Utils
- Implementation:DevExpress Testcafe Runner Fluent API
- Implementation:Lance format Lance RepDef
- Implementation:Openai Openai node Completions Resource
- Implementation:Shiyu coder Kronos Train Model Tokenizer Qlib
- Implementation:Sgl project Sglang Fetch Metrics
- Implementation:Pyro ppl Pyro EscapeMessenger
- Implementation:Datajuicer Data juicer TextEntityDependencyFilter
Heuristics
- Heuristic:NVIDIA TransformerEngine FP8 Recipe Auto Selection
- Heuristic:Vllm project Vllm Attention Backend Selection
- Heuristic:Explodinggradients Ragas Failed Metrics Return NaN
- Heuristic:Sail sg LongSpec Tree Shape Configuration
- Heuristic:Langfuse Langfuse Eval Loop Prevention
- Heuristic:Pola rs Polars Streaming For Large Datasets
- Heuristic:Treeverse LakeFS S3 Multipart Size Constraint
- Heuristic:Wandb Weave Payload Size Limits
- Heuristic:Openai Whisper Compression Ratio Threshold
- Heuristic:Unslothai Unsloth Merge Memory Management
Environments
- Environment:Marker Inc Korea AutoRAG API Keys Configuration
- Environment:Alibaba ROLL Megatron Training Environment
- Environment:Sgl project Sglang Grafana
- Environment:Pytorch Serve Distributed Training Environment
- Environment:Mbzuai oryx Awesome LLM Post training Python Pandas
- Environment:Unstructured IO Unstructured Profiling Tools
- Environment:Kubeflow Pipelines KFP Backend Deployment
- Environment:LMCache LMCache Python Runtime
- Environment:ARISE Initiative Robomimic HDF5 Data Dependencies
- Environment:Guardrails ai Guardrails Python 3 10 Runtime