open-source-models

Tag

Cards List
#open-source-models

@grapeot: Same type of company, three years apart, opposite results. In 2023, Bloomberg trained a 50B model from scratch, fed with 363B tokens of private financial data, official conclusion: not used in any products. In 2026, Thomson Reuters spent $40 million, used...

X AI KOLs Timeline · 2026-08-29 Cached

This article compares the different strategies of Bloomberg and Thomson Reuters in AI model development, analyzes the shift from training large models from scratch to fine-tuning on open bases, and the impact of this trend on vertical AI applications.

0 favorites 0 likes
#open-source-models

@samhogan: introducing fast inference (https://fast.inference.net) fast inference is an LLM API for devs who want to go faster acc…

X AI KOLs Timeline · 2026-08-27 Cached

Fast Inference is an LLM API service that offers fast and affordable access to top open-source and closed-source models for developers, with integrations for coding agents and tools like Claude Code and Codex.

0 favorites 0 likes
#open-source-models

If US labs stay gated and Chinese labs keep shipping open models, the US could hand away the developer ecosystem by accident

Reddit r/ArtificialInteligence · 2026-08-26

The article argues that if US AI labs restrict access to models while Chinese labs release open ones, developers may standardize on Chinese ecosystems, potentially shifting global AI dominance.

0 favorites 0 likes
#open-source-models

@FinanceYF5: Anthropic's outlook appears bleak: Fable 5 is losing market share to more affordable open-source models. According to the Financial Times, in Ramp's enterprise AI spending data, Anthropic's top model Fable 5 only accounts for 11%. The reason is simple: Fa…

X AI KOLs Timeline · 2026-08-26 Cached

According to the Financial Times, Anthropic's strongest model Fable 5 captures only 11% of enterprise AI spending, being overtaken by cheaper open-source models due to its high price.

0 favorites 0 likes
#open-source-models

Are LLMs Safe Beyond Text: Do Emojis Expose Gaps in Safety Evaluation

arXiv cs.CL · 2026-08-20 Cached

This paper investigates the safety of large language models (LLMs) beyond text inputs by examining emoji-augmented prompts, revealing gaps in current safety evaluations and model-dependent vulnerabilities.

0 favorites 0 likes
#open-source-models

Rippling's 2,100 scored runs experiment vs. the Stripe OpenRouter $7B deal

Reddit r/AI_Agents · 2026-08-18

Rippling conducted a benchmark test of 15 AI models on payroll tasks, finding Anthropic's Opus 4.6 performed best but with a 9% failure rate, while Stripe acquired OpenRouter for $7B to help developers choose AI models.

0 favorites 0 likes
#open-source-models

Companies that are adopting AI tend to grow faster - All-in Podcast

Reddit r/ArtificialInteligence · 2026-07-27 Cached

The podcast discussed Palantir's collaboration with Nvidia to develop a sovereign AI operating system, as well as enterprise data sovereignty concerns arising from Anthropic's vertical integration strategy, emphasizing the importance of enterprises using open-source models and private deployment to protect intellectual property and reduce costs.

0 favorites 0 likes
#open-source-models

US threatens sanctions against Chinese AI models over IP theft

TechCrunch AI · 2026-07-21 Cached

US Treasury Secretary Scott Bessent threatens sanctions against Chinese AI models if intellectual property theft is found, escalating the technological competition between US and Chinese AI companies.

0 favorites 0 likes
#open-source-models

More AI Spend Won't Fix Your Supply Chain (4 minute read)

TLDR AI · 2026-07-21 Cached

Pallet launches Custom Models for supply chain teams, enabling enterprise AI sovereignty by training dedicated models on proprietary operational data to improve accuracy, reduce costs, and maintain control.

0 favorites 0 likes
#open-source-models

The real AI race may no longer be at the frontier

TechCrunch AI · 2026-07-14 Cached

The article discusses the growing dominance of open-weight models, especially from Chinese firms, in production AI workloads, challenging the relevance of frontier models from companies like Anthropic and OpenAI.

0 favorites 0 likes
#open-source-models

New method aims to keep kids safe from illegal AI-generated content

MIT News — Artificial Intelligence · 2026-07-13 Cached

MIT researchers developed a technique to audit AI models for their capability to generate child sexual abuse material without producing illegal outputs, achieving 100% accuracy in tests. This method examines hidden model representations to infer whether a model has been fine-tuned for harmful content, providing a scalable way for platforms and law enforcement to detect unsafe models.

0 favorites 0 likes
#open-source-models

GPT-5.6, Grok 4.5, Claude, and Muse Spark build the same 4 apps

Hacker News Top · 2026-07-10 Cached

A detailed comparison of twelve AI models, including GPT-5.6, Grok 4.5, Claude, and open-weight models, tasked with building four different applications across multiple attempts, with all artifacts published for independent evaluation.

0 favorites 0 likes
#open-source-models

@alighodsi: At 11k employees, our AI costs are going up. Which model & harness should we use to lower cost but also retain great qu…

X AI KOLs Timeline · 2026-07-08 Cached

Databricks published an internal benchmark evaluating coding agents on their multi-million line codebase, revealing that harness choice can double cost savings and that open models like GLM 5.2 perform competitively at the highest difficulty levels.

0 favorites 0 likes
#open-source-models

Clouded Judgement 7.3.26 - The End of Compute Scarcity? Not So Fast (14 minute read)

TLDR AI · 2026-07-06 Cached

The article analyzes recent moves by SpaceX and Meta to sell excess AI compute capacity, questioning whether this signals an end to compute scarcity. It argues the deals are short-term and high-priced, and that underlying demand remains strong, refuting the bear thesis.

0 favorites 0 likes
#open-source-models

@Michaelzsguo: If you already have the hardware and just want to try local open models and their agent apps, ODS seems like a quick wa…

X AI KOLs Following · 2026-07-05 Cached

ODS is a full-stack local AI deployment system that helps users run open models and agent apps on their own hardware by calculating compatible models.

0 favorites 0 likes
#open-source-models

Why U.S. Companies Are Quietly Being Run On Chinese AI

Reddit r/ArtificialInteligence · 2026-07-05 Cached

Despite public support for US-made AI, many US tech companies are quietly relying on Chinese open-source models like Qwen and Kimi due to lower cost, higher performance, and faster updates. A USCC report shows 80% of US AI startups use Chinese open-source models as foundations, signaling a significant shift in the infrastructure layer of AI.

0 favorites 0 likes
#open-source-models

Readable but Not Controllable: Neuron-Level Evidence for Medical LLM Hallucination

arXiv cs.CL · 2026-07-02 Cached

This paper investigates whether hallucination in medical LLMs can be detected and controlled at the neuron level. The authors find that while hallucination signals are detectable across many neurons (AUROC 0.77-0.86), they are not easily corrected by steering those same neurons.

0 favorites 0 likes
#open-source-models

End of an Agony. Real production service that uses LLM to earn money my team had made and now we are so happy that it will die. Here are some of my final "experiences".

Reddit r/LocalLLaMA · 2026-07-01

A developer recounts the painful experience of building and eventually shutting down a production LLM-based service for medical appointment scheduling, highlighting issues with model reliability, structured output validation, and provider uptime.

0 favorites 0 likes
#open-source-models

@eglyman: finetuning open source models has been hard to justify. why invest in tuning one when a better model ships within a mon…

X AI KOLs Following · 2026-07-01 Cached

Ethan Glyman announces that his team has developed a method to transfer finetunes between open source models at a fraction of the cost, making finetuning more economical and justifiable.

0 favorites 0 likes
#open-source-models

@quxiaoyin: Turns out Elon is right again. The shittiest layer in AI is the model layer. The real money in AI is in compute, energy…

X AI KOLs Timeline · 2026-07-01

The tweet argues that the AI model layer is the least profitable, while compute, energy, and applications are where the money is, noting that Chinese open-weight models are eroding the margins of companies like OpenAI and Anthropic.

0 favorites 0 likes
Next →
← Back to home

Submit Feedback