Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Mbzuai oryx Awesome LLM Post training Research Trend Analysis
- Workflow:Bitsandbytes foundation Bitsandbytes 8bit Optimizer Training
- Workflow:Openai Openai python Embeddings Generation
- Workflow:Google research Deduplicate text datasets Single file deduplication
- Workflow:Volcengine Verl Supervised Fine Tuning
- Workflow:Lucidrains X transformers DPO Preference Alignment
- Workflow:Haifengl Smile Matrix Decomposition Pipeline
- Workflow:ClickHouse ClickHouse Building From Source
- Workflow:Langchain ai Langchain Streaming Responses
- Workflow:Kubeflow Kubeflow Platform Deployment
Principles
- Principle:Facebookresearch Habitat lab Dataset and Scene Preparation
- Principle:Lm sys FastChat MT Bench Answer Generation
- Principle:Huggingface Diffusers Memory Optimization
- Principle:Langchain ai Langchain Pre Release Validation
- Principle:Apache Shardingsphere Persistence Facade Construction
- Principle:Langchain ai Langchain Compatibility Testing
- Principle:Norrrrrrr lyn WAInjectBench Model Serialization
- Principle:Webdriverio Webdriverio AsyncIterationPattern
- Principle:Nautechsystems Nautilus trader Instrument Registration
- Principle:Interpretml Interpret Feature Group Importance
Implementations
- Implementation:FlagOpen FlagEmbedding EvalDenseRetriever Call
- Implementation:Unslothai Unsloth GEMM Backward Kernels
- Implementation:Pyro ppl Pyro Stats
- Implementation:Haifengl Smile DataFrame Inspection API
- Implementation:Helicone Helicone Mitmproxy Linux
- Implementation:Microsoft Onnxruntime Convert Sklearn
- Implementation:Infiniflow Ragflow Next Request
- Implementation:FlowiseAI Flowise DocumentStoreTable
- Implementation:LaurentMazare Tch rs VarStore Set Kind
- Implementation:Open compass VLMEvalKit CCOCRDataset
Heuristics
- Heuristic:Huggingface Diffusers Memory Offloading Strategy
- Heuristic:Protectai Modelscan Stricter Zip Detection
- Heuristic:Heibaiying BigData Notes Spark Streaming Local Threads Tip
- Heuristic:Huggingface Datatrove Gopher Quality Thresholds
- Heuristic:DevExpress Testcafe Window Resize Correction
- Heuristic:Intel Ipex llm LoRA Target All Linear Layers
- Heuristic:Mlflow Mlflow Model Signature Inference Tips
- Heuristic:BerriAI Litellm Token Counting Buffer
- Heuristic:Run llama Llama index Worker Count Configuration
- Heuristic:Danijar Dreamerv3 XLA GPU Optimization Flags
Environments
- Environment:Astronomer Astronomer cosmos Kubernetes Provider
- Environment:Eric mitchell Direct preference optimization HuggingFace Transformers
- Environment:Openai Openai node Node 20 Runtime
- Environment:PacktPublishing LLM Engineers Handbook VLLM Evaluation Environment
- Environment:Arize ai Phoenix Phoenix Server Runtime
- Environment:Langgenius Dify Vector Database Environment
- Environment:CARLA simulator Carla Python API Runtime
- Environment:Triton inference server Server TRT LLM Deployment
- Environment:Googleapis Python genai Gemini API Key Authentication
- Environment:Guardrails ai Guardrails Python 3 10 Runtime