openai-compatible

Tag

Cards List
#openai-compatible

faaah (Filesystem As An AI Handler)

Lobsters Hottest · 3d ago Cached

FAAAH is a dependency-free, OpenAI-compatible proxy that converts API requests into text files and uses AI coding agents like Claude Code as the backend, enabling reuse of existing subscriptions instead of paying for cloud LLM APIs.

0 favorites 0 likes
#openai-compatible

A unified API for AI model routing (3 minute read)

TLDR AI · 2026-08-05 Cached

Google Cloud API Gateway now offers model routing in Public Preview, providing a serverless ingress layer that accepts OpenAI-compatible requests and dynamically routes them to Gemini, Claude, or OpenAI models.

0 favorites 0 likes
#openai-compatible

@googledevs: No more hardcoded endpoints. Model routing for API Gateway is in Public Preview! Access Gemini, Claude, and OSS models …

X AI KOLs Timeline · 2026-08-04 Cached

Google Cloud API Gateway announces public preview of model routing, letting developers access Gemini, Claude, and OSS models via a single endpoint using OpenAPI specs.

0 favorites 0 likes
#openai-compatible

Introducing celeris-1 (2 minute read)

TLDR AI · 2026-07-27 Cached

Celeris-1 is a new language model using diffusion-based inference architecture, achieving near-GPT-5 level intelligence with 15x faster response times and high throughput.

0 favorites 0 likes
#openai-compatible

Hetzner is working on LLM Inference

Hacker News Top · 2026-07-24 Cached

Hetzner has launched an experimental LLM inference API service, offering an OpenAI-compatible endpoint with the Qwen3.6-35B-A3B-FP8 model. The service is free during the experiment period, has no SLA, and is intended to gather user feedback.

0 favorites 0 likes
#openai-compatible

Ramp Router claims to cut AI costs by up to 30%

Reddit r/ArtificialInteligence · 2026-07-21

Ramp is open-sourcing its internal LLM router that automatically selects the best model for each request to optimize cost and performance.

0 favorites 0 likes
#openai-compatible

@geekbb: 卧槽,兄弟们,CLIProxyAP 上大分,官方认证了属于是,快安装起来吧~ https://github.com/router-for-me/CLIProxyAPI…

X AI KOLs Timeline · 2026-07-12 Cached

CLIProxyAPI 是一个为 CLI 提供兼容 OpenAI/Gemini/Claude/Codex/Grok API 接口的代理服务器,现已支持 OpenAI Codex 和 Claude Code 的 OAuth 访问,方便开发者使用本地或多账户进行 CLI 访问。

0 favorites 0 likes
#openai-compatible

Mesh LLM: distributed AI computing on iroh

Hacker News Top · 2026-07-11 Cached

Mesh LLM is a distributed AI computing platform that pools idle GPUs across multiple machines to run large language models, exposing a single OpenAI-compatible API. It leverages iroh's peer-to-peer networking to enable private, decentralized inference without a central server.

0 favorites 0 likes
#openai-compatible

@Jolyne_AI: Another open-source tool for running AI models locally on GitHub: Shimmy, targeting Ollama's pain points. A single file of only 5MB provides fast, stable local inference with a full OpenAI-compatible API, almost zero integration cost. Built with Rust to maximize performance, it starts in under 100ms and uses about 50MB of memory.

X AI KOLs Timeline · 2026-07-02 Cached

Another open-source tool on GitHub, Shimmy, is a single 5MB file written in Rust that provides fast and stable local inference with a full OpenAI-compatible API, targeting Ollama's pain points. It starts in under 100ms and uses about 50MB of memory.

0 favorites 0 likes
#openai-compatible

@QingQ77: A high-performance API proxy bridging the OpenCode Zen protocol to Anthropic/OpenAI-compatible format, enabling tools like Claude Code and Codex CLI to transparently use domestic models. https://github.com/Kiowx/ope…

X AI KOLs Timeline · 2026-07-02 Cached

opencode-cc is a high-performance Go API proxy that bridges the Anthropic/OpenAI-compatible protocol to the OpenCode Zen protocol, enabling tools such as Claude Code and Codex CLI to transparently use domestic large models including GLM, Kimi, DeepSeek, and Qwen. It supports automatic protocol routing, tool calling, web control panel, and other features.

0 favorites 0 likes
#openai-compatible

I taught myself to code 5 months ago and built an autonomous AI red-team tester — testyourllm.com

Reddit r/artificial · 2026-06-30

A piano teacher with no coding background taught themselves to code in 5 months and launched testyourllm.com, an autonomous AI red-team tester that attacks any OpenAI-compatible LLM endpoint. The attacking AI, Tron, broke Llama 3.3 70B on the first try in live testing.

0 favorites 0 likes
#openai-compatible

Run a vLLM Server on HF Jobs in One Command

Hugging Face Blog · 2026-06-26 Cached

Hugging Face Jobs now allows you to spin up a private OpenAI-compatible LLM endpoint with a single command using vLLM, without provisioning servers or Kubernetes.

0 favorites 0 likes
#openai-compatible

@NFTCPS: Attention freeloaders, an OpenAI-compatible API that aggregates the free quotas of 16 major providers into one – including Google, Groq, Cerebras, Mistral, NVIDIA – totaling roughly 1.7 billion tokens per month, all free. The craziest part is it even…

X AI KOLs Timeline · 2026-06-25 Cached

FreeLLMAPI is an open-source tool that aggregates the free quotas of 16 LLM providers into a single OpenAI-compatible endpoint, with automatic routing and usage tracking, totaling about 1.7 billion tokens per month.

0 favorites 0 likes
#openai-compatible

We Built a Unified API Gateway for AI Agents — Lessons Learned

Reddit r/AI_Agents · 2026-06-22

We built a unified API gateway for AI agents supporting multiple models like Claude, GPT, Codex, and Gemini through a single OpenAI-compatible endpoint. It simplifies integration, billing, and deployment for developers building AI agents and SaaS products.

0 favorites 0 likes
#openai-compatible

@iluciddreaming: GLM 5.2, Kimi K2.7 Code, Step 3.7 Flash all free on ZenMux API. No credit card required, no waitlist. Supports OpenCode, OpenClaw, Cursor, Zed, Hermes, and any Open...

X AI KOLs Timeline · 2026-06-22 Cached

ZenMux API announces free access to multiple models including GLM 5.2, Kimi K2.7 Code, and Step 3.7 Flash, with no credit card or waitlist required. Supports OpenAI-compatible clients such as OpenCode and Cursor.

0 favorites 0 likes
#openai-compatible

I found a secret API that gives $66/week of free GPT-5.5 & Claude Opus credits

Reddit r/artificial · 2026-06-17

FreeModel.dev offers a free API proxy with $66/week in credits for GPT-5.5 and Claude Opus, with referral bonuses.

0 favorites 0 likes
#openai-compatible

@gregbarbosa: Apple didn't, so I did: I made it dead simple to run macOS 27's local and Private Cloud Compute Foundation models in an…

X AI KOLs Following · 2026-06-16 Cached

fm-proxy is a drop-in proxy that lets any app accepting an OpenAI API URL run macOS 27's local and Private Cloud Compute Foundation models, with no extra servers or keys.

0 favorites 0 likes
#openai-compatible

Using Gemma 4 E4B with the LiteRT engine - ~2.4x speedup over Q4 GGUF in text generation, image processing roughly the same

Reddit r/LocalLLaMA · 2026-06-02

A developer benchmarks Gemma 4 E4B using Google's LiteRT engine against a Q4 GGUF quant, finding ~2.4x speedup in text generation due to multi-token prediction (MTP), but only 1.1x in image captioning. The post provides a Python wrapper for an OpenAI-compatible endpoint, though with limitations like deterministic output and single-session engine.

0 favorites 0 likes
#openai-compatible

@gyro_ai: Running large models locally for your own tools involves a mountain of Python dependencies and endless backend configuration — the environment alone scares off many. In reality, most people just want a local interface that works instantly. Shimmy is a Rust-based local inference service, compiled into a single binary, offering an interface identical to OpenAI's…

X AI KOLs Timeline · 2026-05-24 Cached

Shimmy is a lightweight single-binary local inference server that provides a drop-in OpenAI-compatible API for running GGUF models, supporting hot-swapping models and requiring no Python dependencies.

0 favorites 0 likes
#openai-compatible

@DeRonin_: 800M free tokens a month, every major LLM, open source this guy literally made you to forget about any limits repo: htt…

X AI KOLs Following · 2026-05-19 Cached

FreeLLMAPI is an open-source tool that aggregates free tiers from 11 major LLM providers into a single OpenAI-compatible endpoint, routing requests and managing rate limits to deliver ~1B+ tokens per month. It simplifies access to multiple free models through one local server.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback