Large Language Models (LLMs)
Compare the AIGC List products currently published for Large Language Models (LLMs).
Dr7.ai is a medical AI platform offering three specialized models: CheXagent for chest X-ray radiology analysis, BiomedCLIP for medical image-text multimodal understanding, and Clinical Camel for healthcare dialogue. API access is advertised across product pages but the OpenAPI specification endpoint returns 404. A free tier and promotional Pro Plan are available, though standard pricing tiers are not disclosed and no independent benchmarks accompany any model.
Modal is a serverless cloud platform purpose-built for AI and data workloads on GPU infrastructure. Developers write Python, and Modal handles provisioning, scaling, and teardown. It provides an OpenAI-compatible LLM serving layer, container image customization, distributed storage volumes, FastAPI web endpoints, and billing observability. The platform builds its runtime in Rust for memory safety and exposes functionality through a gRPC API and open-source Python client.
RunPod delivers on-demand GPU compute for deploying AI agents, running LLM inference, and orchestrating multi-agent workflows. Serverless endpoints auto-scale and bill only during active compute. Self-hosted inference keeps sensitive data off third-party servers. Free account with no credit card required.
Databricks is a unified data and AI platform built on an open lakehouse architecture available across AWS, Azure, and GCP. It spans data engineering, data science, and AI agent development with native IDE integrations for VS Code and PyCharm, multi-model access, Model Context Protocol (MCP) support, serverless deployment, and enterprise governance through Unity Catalog.
EvoLink is a multi-model AI API platform that gives developers one integration point for language, image, and video models from OpenAI, Anthropic, Google, Midjourney, BytePlus, and xAI. Its Smart Router automates model selection, while transparent per-model pricing and trial credits lower the barrier to multi-provider workflows.
ZeroEntropy is an API platform that combines dense embeddings via zembed-1 with a zerank-1 reranker for production retrieval-augmented generation. It targets regulated enterprise verticals, provides a Python SDK with configurable dimensions and latency modes, and maintains developer community channels. Vendor-published case studies report improved recall and precision in legal, clinical, and customer-support domains, though independent verification is not yet available.
Shusheng AI is a multi-model research platform operated at intern-ai.org.cn, providing centralized access to InternLM (LLM), InternVL (vision-language), a scientific discovery platform, and Fengwu (weather modeling). The platform is positioned for the Chinese academic and research community, with all navigation in Simplified Chinese and minimal public-facing documentation available on the homepage.