qwen-model

Tag

Cards List
#qwen-model

@TeksEdge: A localmaxxer hit ~381 tok/s on a SINGLE RTX 3090 with Qwen3.8-27B. This developer has turned a 24GB RTX 3090 into a mo…

X AI KOLs Timeline · 2d ago Cached

A developer achieved up to 381 tok/s inference speed on a single RTX 3090 with the Qwen3.8-27B model using optimized techniques like DFlash2 and prefix caching, particularly effective for document-based tasks like RAG and coding assistants.

0 favorites 0 likes
#qwen-model

@AlexFinn: This is scary. I downloaded an uncensored version of Qwen 3.8 27B onto my Mac It literally does anything you want. Firs…

X AI KOLs Timeline · 2d ago Cached

The author expresses concern about the ease of accessing uncensored AI models like Qwen 3.8 27B and questions the effectiveness of safety regulations in the face of open source advancements.

0 favorites 0 likes
#qwen-model

Qwen 3.8 is off to College. Got a 34 on the ACT

Reddit r/AI_Agents · 2d ago

An experiment tested the Qwen 3.8 27B AI model on ACT practice exams using vision capabilities, achieving high composite scores of 34-36, showcasing strong performance in standardized testing.

0 favorites 0 likes
#qwen-model

I measured the 3 claims Users in this Sub all handed me on the last local-agent post. One of you out-predicted my own hypothesis. Learn It All not Know It All rules

Reddit r/AI_Agents · 3d ago

The user tested scaling local AI agents with a Qwen 27B model, finding that adding more agents increases throughput only up to a point due to memory bandwidth limits, with long prompts benefiting more from parallelism.

0 favorites 0 likes
#qwen-model

I measured whether 2 local agents hitting 1 model run in parallel or just take turns. Batching is real, but it is not free using QWEN 3.8 27B 4bit on my MacBook Pro M3Max 128 GB Unified Memory 40 Core GPU

Reddit r/AI_Agents · 4d ago

The author experimented with two local agents running in parallel on a MacBook Pro M3Max using the QWEN 3.8 27B 4bit model, finding that batching enables concurrent execution but increases latency, with an optimal agent count around 4.

0 favorites 0 likes
#qwen-model

100$ worth of gpu runs qwen 3.8 27b at 7.39 t/s

Reddit r/LocalLLaMA · 5d ago

A user demonstrates running the Qwen 27b AI model quantized to Q3_K_M on two RX 580 GPUs, achieving 7.39 tokens per second using old DDR3 hardware for under $100.

0 favorites 0 likes
#qwen-model

How many people have 24gb over gpu here?

Reddit r/LocalLLaMA · 6d ago

The author discusses the low adoption of the qwen 3.8 27b model based on download counts and estimates that very few users have the high-VRAM GPUs needed for productive local LLM development.

0 favorites 0 likes
#qwen-model

Which Harness for Local Coding (Qwen 3.8 27b) do you Recommend?

Reddit r/LocalLLaMA · 2026-08-15

A community poll seeking specific harness tool recommendations for local coding with the Qwen 3.8 27b AI model.

0 favorites 0 likes
#qwen-model

Fixed/improved Jinja chat template for Qwen 3.8

Reddit r/LocalLLaMA · 2026-08-14

This article details the improvements to the Jinja chat template for Qwen 3.8 models, correcting issues from previous versions to ensure consistent output quality and benchmark performance.

0 favorites 0 likes
#qwen-model

Towards Inclusive Mobility Modeling: Characterizing and Evaluating Elderly Trajectory Patterns in Urban Systems

arXiv cs.AI · 2026-07-01 Cached

This paper examines how the underrepresentation of elderly riders in mobility datasets introduces systematic bias into mobility modeling, using Citi Bike data from Jersey City. It shows that models trained on majority-dominated populations misrepresent elderly mobility behavior, and that higher-capability models do not necessarily improve subgroup fidelity under limited demographic data.

0 favorites 0 likes
#qwen-model

7900XTX 24GB vram, can finally fit Q6K+MTP with Qwen 3.6 27B at 131k context

Reddit r/LocalLLaMA · 2026-06-20

A guide on optimizing VRAM usage on an AMD 7900XTX to run a 27B Qwen model with Q6K quantization and 131k context by compiling llama.cpp with OpenBLAS and CUDA_FA_ALL_QUANTS, and using kvcache quantization at q5_0/q4_0.

0 favorites 0 likes
#qwen-model

MachinaCheck: Building a Multi-Agent CNC Manufacturability System on AMD MI300X

Hugging Face Blog · 2026-05-10 Cached

MachinaCheck is a multi-agent AI system built on AMD MI300X hardware that automates CNC manufacturability analysis for STEP files using Qwen 2.5 7B models.

0 favorites 0 likes
← Back to home

Submit Feedback