agentic

Tag

Cards List
#agentic

@ollama: Run Ornith with Ollama: ollama run ornith For coding, use it with Claude or Pi: ollama launch claude --model ornith oll…

X AI KOLs Timeline ↗ · 2026-06-27 Cached

Ornith-1.0 is a family of open-source LLMs specialized for agentic coding, available in sizes from 9B to 397B MoE, and can be run via Ollama for use with tools like Claude or Pi.

0 favorites 0 likes
#agentic

@TeksEdge: Been testing Orinth-1.0-35B to see how it stacks up with Qwen3.6-35B over a day's use. Anecdotally, it works as well as…

X AI KOLs Timeline ↗ · 2026-06-27 Cached

A user reports that Ornith-1.0-35B matches Qwen3.6-35B in performance but excels at planning and long task execution, while the developer announces the open-source Ornith-1.0 family of LLMs specialized for agentic coding.

0 favorites 0 likes
#agentic

Previewing GPT‑5.6 Sol: a next-generation model

Hacker News Top ↗ · 2026-06-26 Cached

OpenAI previews the GPT-5.6 series including flagship Sol, balanced Terra, and affordable Luna models, featuring improved reasoning, agentic capabilities, and robust safety measures, with a limited preview before broader availability.

0 favorites 0 likes
#agentic

Dockerless: Environment-Free Program Verifier for Coding Agents

Hugging Face Daily Papers ↗ · 2026-06-26 Cached

This paper introduces Dockerless, an environment-free agentic patch verifier that evaluates code patches without execution, outperforming existing open-source verifiers and enabling efficient post-training for coding agents.

0 favorites 0 likes
#agentic

@MaxForAI: hey! There's actually such an amazing model?? Ornith @ornith_ just open-sourced a new open-source model family. Includes 9B Dense, 31B Dense, 35B MoE, and 397B MoE. And the official benchmarks are impressive, even comparable to GLM…

X AI KOLs Timeline ↗ · 2026-06-25 Cached

Ornith has open-sourced the Ornith-1.0 model family, which includes multiple sizes such as 9B Dense, 31B Dense, 35B MoE, and 397B MoE. It achieves leading performance on coding tasks, even rivaling GLM5.2.

0 favorites 0 likes
#agentic

@anvie: Tested Ornith-1.0-9B, and its impressive for a model of that size. I don't believe this is just 9B!

X AI KOLs Following ↗ · 2026-06-25 Cached

Ornith-1.0 is a family of open-source LLMs specialized for agentic coding, spanning sizes from 9B to 397B and achieving state-of-the-art performance among open-source models of comparable size.

0 favorites 0 likes
#agentic

@rohanpaul_ai: Another fantastic open source release. DeepReinforce just dropped Ornith-1.0, an MIT-licensed open-source family of age…

X AI KOLs Timeline ↗ · 2026-06-25 Cached

DeepReinforce releases Ornith-1.0, an MIT-licensed open-source family of agentic coding LLMs including a 397B MoE model that surpasses Claude Opus 4.7 on SWE-Bench and Terminal-Bench, using a novel self-improving training strategy.

0 favorites 0 likes
#agentic

@liquidai: Introducing LFM2.5-230M: our smallest model yet, built to run fast anywhere (CPUs, NPUs, and GPUs) to enable agentic ta…

X AI KOLs Timeline ↗ · 2026-06-25 Cached

Liquid AI releases LFM2.5-230M, a small 230M parameter model optimized for fast inference on CPUs, NPUs, and GPUs, targeting agentic tasks on devices like phones and robots.

0 favorites 0 likes
#agentic

Introducing computer use in Gemini 3.5 Flash

Google DeepMind Blog ↗ · 2026-06-24 Cached

Gemini 3.5 Flash now natively supports computer use as a built-in tool, enabling developers to build agents that can interact across browser, mobile, and desktop environments for long-horizon automation tasks like software testing and knowledge work.

0 favorites 0 likes
#agentic

Pulse

Product Hunt ↗ · 2026-06-24

Pulse is a permission-aware, proactive, and agentic AI brain for companies, launched on Product Hunt.

0 favorites 0 likes
#agentic

llama.cpp's web UI now supports executing model generated JavaScript in the browser, through Web Workers (opt in)

Reddit r/LocalLLaMA ↗ · 2026-06-24

llama.cpp's web UI now supports executing model-generated JavaScript in a sandboxed iframe via Web Workers, enabling lightweight agentic code execution as an opt-in feature.

0 favorites 0 likes
#agentic

@_TobiasLee: Seed 2.1 from Bytedance achieved impressive results on two of our benchmarks. Claw-Eval (Multimodal, https://claw-eval.…

X AI KOLs Timeline ↗ · 2026-06-24 Cached

ByteDance's Seed 2.1 model achieved strong results on multimodal agentic (Claw-Eval) and long video understanding (Video-MME) benchmarks, though a gap remains between perception and agentic capabilities.

0 favorites 0 likes
#agentic

Autodata: An agentic data scientist to create high quality synthetic data

Hugging Face Daily Papers ↗ · 2026-06-24 Cached

Autodata is a method that enables AI agents to act as data scientists to create high-quality synthetic training data through meta-optimization, achieving improved performance across computer science, legal reasoning, and mathematical tasks.

0 favorites 0 likes
#agentic

@sashimikun_void: Another day watching agentic Slack startups get wiped. I've heard founders say, "Don't worry, they won't have time to f…

X AI KOLs Following ↗ · 2026-06-23 Cached

A developer reflects on how AI agents are eliminating Slack startup niches, while ClaudeDevs reveals that Claude Code now writes 65% of their product team's code, including the Claude Tag tool itself.

0 favorites 0 likes
#agentic

@analogalok: gemma-4-12B-agentic-fable5-composer2.5 V2 is out. the agentic upgrade to the model trained on Fable 5's reasoning. Runn…

X AI KOLs Timeline ↗ · 2026-06-21 Cached

A new fine-tuned version of Gemma 4 12B, trained on Fable 5's reasoning, achieves a significant jump in agentic coding benchmarks (from 15% to 55%) and can run locally on an 8GB VRAM GPU using a custom fork of llama.cpp.

0 favorites 0 likes
#agentic

@maximelabonne: LFM2.5-ColBERT-350M is a surprisingly reliable smart tool selector. We gave it 151 tools, and it consistently surfaces …

X AI KOLs Following ↗ · 2026-06-18 Cached

LFM2.5-ColBERT-350M is a model that reliably selects the most relevant tools from a set of 151, saving tokens and improving accuracy, ideal for agentic edge models.

0 favorites 0 likes
#agentic

Git platform built for agentic era

Hacker News Top ↗ · 2026-06-18

A new git platform designed for the agentic era, likely targeting AI-driven development workflows.

0 favorites 0 likes
#agentic

@Saboo_Shubham_: Turn your AI coding agent into an agentic video production studio. 100% Opensource.

X AI KOLs Following ↗ · 2026-06-18 Cached

A 100% opensource tool that transforms an AI coding agent into an agentic video production studio, announced by @Saboo_Shubham_.

0 favorites 0 likes
#agentic

Is it agentic enough? Benchmarking open models on your own tooling

Hugging Face Blog ↗ · 2026-06-18 Cached

This blog post introduces a benchmark methodology for evaluating how well open models perform on agentic coding tasks, focusing not just on accuracy but on the efficiency of the agent's process. It provides a customizable tooling harness using the pi coding agent and tests across models and library revisions.

0 favorites 0 likes
#agentic

@witcheer: this is the first Qwen3.6-27B coding tune I've measured that improves real bug-fixing (!!!). - quality (MMLU/ARC/HellaS…

X AI KOLs Timeline ↗ · 2026-06-17 Cached

A community fine-tune of Qwen3.6-27B improves real bug-fixing on SWE-bench while maintaining quality, unlike synthetic distillations that regress.

0 favorites 0 likes
← Previous
Next →
← Back to home

Submit Feedback