🛠️

Developer Tools & MLOps

Vector databases, LLM observability, model evaluation suites, fine-tuning platforms, and prompts ops.

9 tools in this category Filter in global directory ➔
C

Cohere

by Cohere
freemium
🛠️ Developer Tools & MLOps

Enterprise AI platform specializing in multilingual embeddings, reranking, and enterprise search.

  • Cohere Rerank 3 model for boosting RAG search accuracy
  • Embed v3 multilingual vector embedding models
  • Deployable on AWS SageMaker, Azure, and private VPCs
G

Groq LPU

by Groq, Inc.
freemium
🛠️ Developer Tools & MLOps

Ultra-high-speed AI inference engine powered by custom LPU (Language Processing Unit) hardware.

  • Ultra-fast 500+ tokens/second LLM generation speed
  • Custom LPU (Language Processing Unit) chip architecture
  • OpenAI-compatible REST API for easy drop-in replacement
H

Hugging Face

by Hugging Face, Inc.
freemium
🛠️ Developer Tools & MLOps

The central open-source community platform for hosting machine learning models, datasets, and AI apps.

  • 500,000+ open-source ML models and datasets
  • Hugging Face Spaces for hosting interactive web demos
  • Inference Endpoints for managed cloud model deployment
L

LangSmith

by LangChain
freemium
🛠️ Developer Tools & MLOps

LLM application observability, tracing, evaluation, and prompt engineering platform.

  • Full execution chain visual trace logging (spans, inputs, outputs, token costs)
  • Automated evaluation suites and LLM-as-a-judge scoring frameworks
  • Prompt Playground for testing prompt variations against production traces
O

Ollama

by Ollama
free
🛠️ Developer Tools & MLOps

Open-source tool for running large language models locally on macOS, Linux, and Windows.

  • Simple terminal CLI for model downloading and execution
  • Built-in local REST API compatible with OpenAI endpoint format
  • Supports Llama 3, DeepSeek-R1, Mistral, Gemma, and Phi
P

Pinecone

by Pinecone Systems
freemium
🛠️ Developer Tools & MLOps

Fully-managed serverless vector database engineered for ultra-low latency RAG search.

  • Serverless vector architecture with pay-per-query consumption pricing
  • Sub-100 millisecond vector search latency at scale
  • Metadata filtering for combining structured database attributes with semantic vectors
V

vLLM

by vLLM Project
free
🛠️ Developer Tools & MLOps

High-throughput and memory-efficient open-source LLM serving engine powered by PagedAttention.

  • PagedAttention algorithm for optimized KV cache memory management
  • OpenAI-compatible API server endpoint
  • Supports continuous batching and multi-GPU tensor parallelism
W

Weaviate

by Weaviate B.V.
freemium
🛠️ Developer Tools & MLOps

Open-source multi-modal vector database featuring native vectorization and hybrid search.

  • Open-source core engine with local Docker/Kubernetes deployment capability
  • Hybrid Search engine combining BM25 keyword matching and dense vector search
  • Native module integration for automated vectorization (OpenAI, HuggingFace, Cohere)
W

Weights & Biases

by Weights & Biases, Inc.
freemium
🛠️ Developer Tools & MLOps

Developer platform for MLOps, model training tracking, dataset versioning, and LLM evaluation.

  • Experiment tracking for PyTorch, TensorFlow, and Hugging Face
  • Dataset and model artifact version control
  • W&B Prompts for LLM trace evaluation