Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Apache Spark Application Submission
- Workflow:Langfuse Langfuse Dataset experiment pipeline
- Workflow:Huggingface Datasets Dataset Loading and Exploration
- Workflow:Guardrails ai Guardrails Custom Validator Development
- Workflow:Online ml River Online Clustering
- Workflow:Haotian liu LLaVA Benchmark Evaluation
- Workflow:Fede1024 Rust rdkafka Mock Cluster Testing
- Workflow:Explodinggradients Ragas LLM Benchmarking
- Workflow:Huggingface Datatrove Minhash Deduplication
- Workflow:Dagster io Dagster DSPy Optimization
Principles
- Principle:NVIDIA TransformerEngine HF Decoder Layer Replacement
- Principle:Tencent Ncnn Bounding Box Decoding
- Principle:Fastai Fastbook Numericalization
- Principle:Groq Groq python File Upload
- Principle:Anthropics Anthropic sdk python Tool Execution
- Principle:Langchain ai Langgraph Tool Definition
- Principle:Langgenius Dify Deployment Upgrade
- Principle:Unslothai Unsloth Supervised Finetuning
- Principle:Arize ai Phoenix Batch Span Annotation
- Principle:Mlflow Mlflow Model Version Management
Implementations
- Implementation:Tensorflow Serving Observer
- Implementation:Microsoft DeepSpeedExamples DeepSpeed Save Checkpoint
- Implementation:Datajuicer Data juicer LLMQualityScoreFilter
- Implementation:Ollama Ollama Agent Approval
- Implementation:Isaac sim IsaacGymEnvs RLGPUAlgoObserver Metrics
- Implementation:Online ml River Cluster ODAC
- Implementation:Huggingface Datatrove ParquetReader
- Implementation:Open compass VLMEvalKit GSM8K V Utils
- Implementation:Predibase Lorax LLaVA NeXT Model
- Implementation:Haosulab ManiSkill BaseDigitalTwinEnv
Heuristics
- Heuristic:Langgenius Dify Token Refresh Loop Prevention
- Heuristic:Haosulab ManiSkill Num Envs Backend Selection
- Heuristic:Arize ai Phoenix AIMD Concurrency Control
- Heuristic:SeleniumHQ Selenium Bazel Hermetic Build Requirement
- Heuristic:Microsoft DeepSpeedExamples RLHF Hyperparameter Guide
- Heuristic:Mage ai Mage ai Batch Size Tuning
- Heuristic:Princeton nlp Tree of thought llm Ad Hoc Value Map Scoring
- Heuristic:Cypress io Cypress Global Install Warning
- Heuristic:Allenai Open instruct Logprob Clamping
- Heuristic:Intel Ipex llm NF4 Quantization Best Practice
Environments
- Environment:Sdv dev SDV Python Runtime
- Environment:Arize ai Phoenix OpenTelemetry SDK
- Environment:Openai Evals Python Runtime
- Environment:Liu00222 Open Prompt Injection CUDA Environment
- Environment:ARISE Initiative Robomimic HuggingFace Hub Dependencies
- Environment:Ggml org Llama cpp CMake Build Environment
- Environment:Deepspeedai DeepSpeed Multi Accelerator Environment
- Environment:Axolotl ai cloud Axolotl Python Runtime
- Environment:Turboderp org Exllamav2 Build Toolchain
- Environment:Shiyu coder Kronos Qlib Data Environment