Main Page
Welcome to Leeroopedia
Your ML & Data Knowledge Wiki. Best practices and expert-level knowledge for Machine Learning and Data Engineering, covering 1000+ frameworks and libraries from training to deployment.
Browse implementation patterns, configuration guides, debugging heuristics, and battle-tested defaults for frameworks like vLLM, DeepSpeed, Megatron-LM, FlashAttention, Triton, Unsloth, LangChain, and many more. Every page is structured so both humans and AI agents can find what they need fast.
Connect your AI coding agent. Plug Leeroopedia into your favorite coding agent, and let it build robust AI/ML systems autonomously:
- SuperML plugin — converts your AI coding agent into an expert ML engineer with agentic memory
- Leeroopedia MCP — search over best-practices and skills of ML/AI
- Kapso — experimentation platform for autonomous AI/ML software building
Browse by Category
| Category | Description | Browse |
|---|---|---|
| Workflows | Step-by-step processes and procedures | Browse All |
| Principles | Core ideas and foundational knowledge | Browse All |
| Implementations | Code-level details and modules | Browse All |
| Heuristics | Best practices and guidelines | Browse All |
| Environments | Setup and configuration guides | Browse All |
Explore Pages
Workflows
- Workflow:Cypress io Cypress Component Test Execution
- Workflow:Google research Deduplicate text datasets Single file deduplication
- Workflow:Speechbrain Speechbrain Text to Speech Training
- Workflow:Fede1024 Rust rdkafka Async Stream Processing
- Workflow:Mlc ai Web llm Text Embeddings And RAG
- Workflow:Dagster io Dagster RAG Pipeline
- Workflow:Huggingface Datatrove Common Crawl Processing
- Workflow:Kubeflow Kubeflow Release Management
- Workflow:Nautechsystems Nautilus trader Strategy development
- Workflow:Neuml Txtai RAG Pipeline
Principles
- Principle:Huggingface Datatrove Disk Writing Framework
- Principle:Helicone Helicone Result Type Pattern
- Principle:Huggingface Datatrove Parquet Data Reading
- Principle:Getgauge Taiko Request Redirect
- Principle:Microsoft Playwright Assertion and Verification
- Principle:Langchain ai Langgraph Tool Validation
- Principle:Triton inference server Server Health Check API
- Principle:Huggingface Alignment handbook QLoRA Quantized Finetuning
- Principle:Huggingface Datasets Environment Reporting
- Principle:Snorkel team Snorkel Multitask Classifier Construction
Implementations
- Implementation:LMCache LMCache VLLM Serve Prefiller
- Implementation:Evidentlyai Evidently Legacy IsValidJSON Feature
- Implementation:Apache Kafka TopicCommand Execute
- Implementation:Puppeteer Puppeteer Viewport
- Implementation:Ggml org Ggml Sycl softmax
- Implementation:PacktPublishing LLM Engineers Handbook HfApi Model Info
- Implementation:Puppeteer Puppeteer NgSchematics Builder
- Implementation:Evidentlyai Evidently Legacy Classification Calculations
- Implementation:DistrictDataLabs Yellowbrick Sphinx Documentation Config
- Implementation:Datajuicer Data juicer FlaggedWordFilter
Heuristics
- Heuristic:Roboflow Rf detr Layer Wise LR Decay
- Heuristic:Sdv dev SDV Sampling Retry Tuning
- Heuristic:Spotify Luigi Marker Table Idempotency
- Heuristic:Turboderp org Exllamav2 Quantization Conversion Tips
- Heuristic:Allenai Open instruct BFloat16 Training
- Heuristic:Bitsandbytes foundation Bitsandbytes Outlier Threshold Detection
- Heuristic:Pola rs Polars Collect All For Diverging Queries
- Heuristic:Apache Paimon Vector Index Configuration Tips
- Heuristic:Sail sg LongSpec Tree Shape Configuration
- Heuristic:Fede1024 Rust rdkafka Producer Flush Before Drop
Environments
- Environment:Treeverse LakeFS Web UI Environment
- Environment:Rapidsai Cuml CUDA GPU
- Environment:EvolvingLMMs Lab Lmms eval Python Runtime Environment
- Environment:Vllm project Vllm Buildkite
- Environment:Huggingface Datatrove S3 Storage Environment
- Environment:Microsoft Playwright Browser Binaries Environment
- Environment:Nightwatchjs Nightwatch Android Mobile Testing
- Environment:OpenHands OpenHands Integration Credentials
- Environment:Triton inference server Server TRT LLM Deployment
- Environment:Mlflow Mlflow OpenAI LLM Integration Environment