NVIDIA NIM
NVIDIA operates in accelerated computing, focusing on artificial intelligence, high-performance computing, and graphics processing. The company provides products including graphics processing units (GPUs), AI platforms, and data center infrastructure for various applications. NV…
Fireworks.ai
Fireworks AI specializes in generative artificial intelligence platform services, focusing on inference and model fine-tuning within the artificial intelligence sector. The company offers an inference engine for building production-ready AI systems and provides a serverless depl…
Together
Together AI focuses on the development, training, fine-tuning, and deployment of generative artificial intelligence (AI) models. The company provides services including AI model training, inference, and utilizes a cloud-based infrastructure. Together AI serves various sectors by…
NVIDIA FasterTransformer
Transformer related optimization, including BERT, GPT
SambaNova
SambaNova Systems specializes in enterprise-scale artificial intelligence (AI) platforms and operates within the artificial intelligence and machine learning sectors. The company offers products designed for the development, training, and deployment of AI models, including gener…
AWS Bedrock
Amazon (NASDAQ: AMZN) is a technology and e-commerce company that operates across various sectors, including online retail and cloud computing. The company provides products and services, including an online marketplace for consumers and businesses, cloud computing solutions, an…
Clarifai
Clarifai focuses on artificial intelligence (AI), specializing in computer vision, natural language processing, and audio recognition across various sectors. The company provides an AI lifecycle platform for building, training, and deploying AI models, including data labeling an…
TensorFlow
An Open Source Machine Learning Framework for Everyone
ONNX Runtime
Open standard for machine learning interoperability
Triton Inference Server
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
Hugging Face TGI
Large Language Model Text Generation Inference
vLLM
A high-throughput and memory-efficient inference and serving engine for LLMs