Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Shiyu coder Kronos Single Series Prediction
- Workflow:Junyanz Pytorch CycleGAN and pix2pix Pretrained Inference
- Workflow:Duckdb Duckdb Code Generation Pipeline
- Workflow:Onnx Onnx Reference Evaluation
- Workflow:Farama Foundation Gymnasium Vectorized Environment Training
- Workflow:BerriAI Litellm SDK Completion
- Workflow:Danijar Dreamerv3 Single Process Training
- Workflow:Dotnet Machinelearning Time Series Forecasting
- Workflow:Danijar Dreamerv3 Distributed Parallel Training
- Workflow:Wandb Weave Tracing Setup
Principles
- Principle:Marker Inc Korea AutoRAG Strategy Selection
- Principle:ClickHouse ClickHouse Network Interface Enumeration
- Principle:Huggingface Trl DPO Preference Dataset Loading
- Principle:OWASP Www project top 10 for large language model applications Style Guide Conformance
- Principle:Openclaw Openclaw Gateway Server Startup
- Principle:Ggml org Ggml SYCL GPU Computation
- Principle:Langchain ai Langgraph Graph Instantiation
- Principle:Webdriverio Webdriverio CapabilityNormalization
- Principle:Spcl Graph of thoughts Ground Truth Evaluation
- Principle:NVIDIA NeMo Curator Text Cleaning and Normalization
Implementations
- Implementation:Sktime Pytorch forecasting TimeSeriesDataSet To Dataloader
- Implementation:BerriAI Litellm Custom Secret Manager Loader
- Implementation:ArroyoSystems Arroyo Checkpoint Api Types
- Implementation:Infiniflow Ragflow FileUploader Component
- Implementation:Treeverse LakeFS Java SDK Model RefList
- Implementation:Open compass VLMEvalKit Infer Data Job
- Implementation:Open compass VLMEvalKit MathCanvas Utils
- Implementation:Datajuicer Data juicer VideoMotionScorePtlflowFilter
- Implementation:Speechbrain Speechbrain Train TimersAndSuch Multistage
- Implementation:Apache Paimon Blob From Descriptor
Heuristics
- Heuristic:Google deepmind Mujoco Mesh Quality For Collision
- Heuristic:Ucbepic Docetl Rate Limit Exponential Backoff
- Heuristic:Allenai Open instruct BFloat16 Training
- Heuristic:Datajuicer Data juicer Batch Size Adaptation
- Heuristic:Mbzuai oryx Awesome LLM Post training Paper Deduplication Via Dict
- Heuristic:Vibrantlabsai Ragas Reasoning Model Parameter Constraints
- Heuristic:Mistralai Client python Resource Context Manager
- Heuristic:ChenghaoMou Text dedup SimHash Optimization Ceiling
- Heuristic:LaurentMazare Tch rs Safetensors Format Preference
- Heuristic:LaurentMazare Tch rs Hidden Dimension Alignment
Environments
- Environment:Volcengine Verl SGLang Rollout Environment
- Environment:Tencent Ncnn Build Environment
- Environment:Huggingface Peft BitsAndBytes Quantization
- Environment:VainF Torch Pruning PyTorch Python Core
- Environment:Alibaba ROLL SGLang Inference Environment
- Environment:Mlc ai Mlc llm WebGPU Browser Environment
- Environment:Microsoft DeepSpeedExamples CIFAR10 Training Environment
- Environment:Onnx Onnx Python Runtime Environment
- Environment:FMInference FlexLLMGen CUDA GPU
- Environment:Datajuicer Data juicer Ray Cluster Environment