LongCat-2.0 is a LongCat AI model announcement highlighting a 1.6 trillion-parameter system trained entirely on domestic chips. The available pages confirm the product’s scale and a separate pricing page, but not usage details or plan structure.
TuneLLM is an enterprise platform that distills recurring Claude- or GPT-style workflows into smaller fine-tuned models inside your infrastructure. It is aimed at teams that want benchmarked quality on narrow LLM tasks at lower inference cost.
Alvoff Inference is an OpenAI-compatible API for speech-to-text, text-to-speech, embeddings, and chat/code generation. It is built for developers who want to swap in a different base URL, use familiar SDKs, and pay per request.
RunInfra helps teams turn open-source models into production inference stacks by benchmarking GPUs, tuning supported runtime paths, and either deploying a managed API or exporting the stack for self-hosting.
ClinePass is a paid subscription offer for accessing curated open weight models in Cline, with a Product Hunt first-month promotion. It is aimed at developers who want a simpler setup for IDE and CLI coding workflows.
discode.ai is a browser-based AI chat product that routes prompts across many models and adds controls for eco impact, local privacy, and multi-model verification. It helps users choose how each answer should balance cost, confidentiality, and confidence.
Heron is a passive observability tool for AI agents and LLM APIs. It reconstructs agent turns, tool calls, and LLM interactions from network traffic without requiring SDK changes or an in-path proxy.
Oxlo.ai is an AI inference API with OpenAI-compatible access and request-based monthly pricing. It is designed for developers and AI teams that want predictable costs for assistants, document workflows, and other production inference workloads.
Crewdle Chat brings GPT, Claude, Gemini, and Grok into one chat workspace for business teams. It supports web search, uploaded documents, and token-based billing with no per-seat fees.
TruthAgent is a web app that runs a question through multiple AI models and surfaces consensus, disagreements, and confidence. It offers a free tier plus credit-based Pro and Pro+ plans for deeper research and decision support.
Sakana Fugu is a multi-agent model API that routes tasks across specialized models through one OpenAI-compatible endpoint. It is positioned for coding, reasoning, and other quality-critical workflows, with configurable agent selection and no public pricing shown in the provided sources.
Second Hand Tokens is an AI token marketplace and API gateway that lets developers buy unused model credits at 50% off retail. It also supports selling unused credits through the same proxied workflow.
co/core is a cooperative for AI inference that pools member-owned Macs to run open models. It supports OpenAI-compatible clients and lets Mac owners contribute compute through a macOS app.
Poolside is a foundation model company for enterprise software work, with agents, developer surfaces, and in-environment deployment options. It is aimed at teams that need governed AI workflows inside their own boundaries.
IN THE WEIGHTS is a web app that checks how strongly leading AI models recognize a name, then presents the result as a strength score, leaderboard placement, and per-name page. It is useful for exploring whether a person or character is represented in model weights without using web search.
Mellum is JetBrains’ open-source family of language models for real-world AI workloads, with a focus on low-latency inference and coding-oriented tasks. It is positioned for development workflows that combine code, context, and natural language.
GLM-5.2 is Z.ai’s flagship long-horizon model for coding and agent workflows, with a 1M-token context and effort-level controls. It is available through Z.ai’s API platform and coding plan for tools such as Claude Code, Cline, OpenCode, and Clawdbot/OpenClaw.
Wolfram Language 15 adds notebook-based AI assistance, programmatic LLM integration, and updated computational, interface, and tooling features across technical workflows.
LLM Gateway is a web-based AI playground for chatting with 210+ models and using image, video, audio, and Canvas workflows from one account. It offers monthly plans, pay-as-you-go credits, and a free start option.
Xiaomi MiMo is an AI model platform with a browser demo and API access. The homepage introduces the MiMo family and links to technical posts, but does not publish detailed pricing or integration information in the available source.