openai-compatible

Tag

Cards List
#openai-compatible

same question, different words, you still pay twice: i built a proxy for that

Reddit r/AI_Agents ↗ · yesterday

MorrowCache is a local OpenAI-compatible proxy that caches AI model responses based on intent similarity to reduce latency and token costs, using an adjudicator to check for cached answers.

0 favorites 0 likes
#openai-compatible

2BA.AI

Product Hunt ↗ · 2026-09-16 Cached

2BA.AI provides EU-hosted AI infrastructure with a flat €20/month fee, offering 4,500 requests per 5-hour window and integration with tools like Cursor and VS Code while ensuring GDPR compliance.

0 favorites 0 likes
#openai-compatible

I built a serverless hosting platform for LoRA adapters with vLLM

Reddit r/LocalLLaMA ↗ · 2026-09-13

Lorivo is a serverless platform that allows sharing GPU servers for hosting multiple LoRA adapters, simplifying deployment and reducing costs for fine-tuned AI models.

0 favorites 0 likes
#openai-compatible

@EmperoAI: thank you for 2000 follower! we have hosted Qwen3.8-27B for you for free! https://free.empero.org/v1 use any api key!

X AI KOLs Timeline ↗ · 2026-08-23 Cached

EmperoAI celebrates reaching 2000 followers by launching a free community API endpoint for the Qwen3.8-27B-FP8 AI model, allowing developers to access it with any API key.

0 favorites 0 likes
#openai-compatible

@smark0428: Folks, DeepSeek V4 Pro is now basically free! TeamoRouter, as a unified entry point for AI models and agents, lets you get the following at virtually 0 cost (at 0.1% of the price): deepseek v4 pro 0813 official version—free, deepseek v…

X AI KOLs Timeline ↗ · 2026-08-21 Cached

TeamoRouter offers free or low-cost access to AI models like DeepSeek V4 Pro, compatible with development tools such as Claude Code and Codex, providing convenient integration through a unified API.

0 favorites 0 likes
#openai-compatible

@ericzakariasson: if you like grok 4.6, you can also use it in the api! it's an openai compatible format, so you can just use it as a dro…

X AI KOLs Following ↗ · 2026-08-20 Cached

Eric Zakariasson announces that Grok 4.6 is available via an API compatible with OpenAI's format, enabling easy integration as a drop-in replacement.

0 favorites 0 likes
#openai-compatible

@JustinGorya: Hetzner now gives you FREE access to Qwen 3.8 27B This model is currently on same level as GPT-5.6 Luna In my testing i…

X AI KOLs Timeline ↗ · 2026-08-18 Cached

Hetzner offers free access to the Qwen 3.8 27B AI model via an OpenAI-compatible API, claiming it performs better than GPT-5.6 Luna in tests.

0 favorites 0 likes
#openai-compatible

Full 1M context V4-Flash without owning eight GPUs

Reddit r/ArtificialInteligence ↗ · 2026-08-16

The article introduces Gonka, a decentralized inference network that enables access to the V4-Flash AI model with full 1M context without requiring local GPU ownership, using an OpenAI-compatible interface.

0 favorites 0 likes
#openai-compatible

CORS Chat

Simon Willison's Blog ↗ · 2026-08-15 Cached

CORS Chat is a browser-based tool for chatting with OpenAI Responses-compatible API endpoints, featuring custom headers, local conversation saving, and progressive SVG rendering.

0 favorites 0 likes
#openai-compatible

@victormustar: I deployed FREE public endpoint for Qwen3.8-27B no token needed, OpenAI-compatible, light rate limiting. Powered by Hug…

X AI KOLs Following ↗ · 2026-08-15 Cached

A free public endpoint for the Qwen3.8-27B AI model has been deployed, offering an OpenAI-compatible API with vision support, tool calls, and a large context window, powered by Hugging Face Inference Endpoints for at least 72 hours.

0 favorites 0 likes
#openai-compatible

DeepSeek API Pricing Update

Hacker News Top ↗ · 2026-08-13 Cached

DeepSeek launches V4-Pro and V4-Flash with flexible reasoning effort, native OpenAI Responses API support, and optimized agent workflows for Codex, available via API and app/web.

0 favorites 0 likes
#openai-compatible

faaah (Filesystem As An AI Handler)

Lobsters Hottest ↗ · 2026-08-08 Cached

FAAAH is a dependency-free, OpenAI-compatible proxy that converts API requests into text files and uses AI coding agents like Claude Code as the backend, enabling reuse of existing subscriptions instead of paying for cloud LLM APIs.

0 favorites 0 likes
#openai-compatible

A unified API for AI model routing (3 minute read)

TLDR AI ↗ · 2026-08-05 Cached

Google Cloud API Gateway now offers model routing in Public Preview, providing a serverless ingress layer that accepts OpenAI-compatible requests and dynamically routes them to Gemini, Claude, or OpenAI models.

0 favorites 0 likes
#openai-compatible

@googledevs: No more hardcoded endpoints. Model routing for API Gateway is in Public Preview! Access Gemini, Claude, and OSS models …

X AI KOLs Timeline ↗ · 2026-08-04 Cached

Google Cloud API Gateway announces public preview of model routing, letting developers access Gemini, Claude, and OSS models via a single endpoint using OpenAPI specs.

0 favorites 0 likes
#openai-compatible

Introducing celeris-1 (2 minute read)

TLDR AI ↗ · 2026-07-27 Cached

Celeris-1 is a new language model using diffusion-based inference architecture, achieving near-GPT-5 level intelligence with 15x faster response times and high throughput.

0 favorites 0 likes
#openai-compatible

Hetzner is working on LLM Inference

Hacker News Top ↗ · 2026-07-24 Cached

Hetzner has launched an experimental LLM inference API service, offering an OpenAI-compatible endpoint with the Qwen3.6-35B-A3B-FP8 model. The service is free during the experiment period, has no SLA, and is intended to gather user feedback.

0 favorites 0 likes
#openai-compatible

Ramp Router claims to cut AI costs by up to 30%

Reddit r/ArtificialInteligence ↗ · 2026-07-21

Ramp is open-sourcing its internal LLM router that automatically selects the best model for each request to optimize cost and performance.

0 favorites 0 likes
#openai-compatible

@geekbb: 卧槽,兄弟们,CLIProxyAP 上大分,官方认证了属于是,快安装起来吧~ https://github.com/router-for-me/CLIProxyAPI…

X AI KOLs Timeline ↗ · 2026-07-12 Cached

CLIProxyAPI 是一个为 CLI 提供兼容 OpenAI/Gemini/Claude/Codex/Grok API 接口的代理服务器,现已支持 OpenAI Codex 和 Claude Code 的 OAuth 访问,方便开发者使用本地或多账户进行 CLI 访问。

0 favorites 0 likes
#openai-compatible

Mesh LLM: distributed AI computing on iroh

Hacker News Top ↗ · 2026-07-11 Cached

Mesh LLM is a distributed AI computing platform that pools idle GPUs across multiple machines to run large language models, exposing a single OpenAI-compatible API. It leverages iroh's peer-to-peer networking to enable private, decentralized inference without a central server.

0 favorites 0 likes
#openai-compatible

@Jolyne_AI: Another open-source tool for running AI models locally on GitHub: Shimmy, targeting Ollama's pain points. A single file of only 5MB provides fast, stable local inference with a full OpenAI-compatible API, almost zero integration cost. Built with Rust to maximize performance, it starts in under 100ms and uses about 50MB of memory.

X AI KOLs Timeline ↗ · 2026-07-02 Cached

Another open-source tool on GitHub, Shimmy, is a single 5MB file written in Rust that provides fast and stable local inference with a full OpenAI-compatible API, targeting Ollama's pain points. It starts in under 100ms and uses about 50MB of memory.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback