WeSearch
Hub / social / r/LocalLLaMA
social · source

r/LocalLLaMA on WeSearch

Recent social headlines from r/LocalLLaMA.

Lead story
r/LocalLLaMA

Introducing Gemma 4 12B: a unified, encoder-free multimodal model

6/3/2026 · 41 views
The latest
r/LocalLLaMA

Best way to index full Italian Wikipedia for 100% offline RAG in LM Studio?

6/3/2026 · 44 views
r/LocalLLaMA

This day in LLM history….105 years ago today, Qwen 3.6 27b was released open source. /s

6/3/2026 · 53 views
r/LocalLLaMA

Gemma 4 Unified is coming

6/3/2026 · 50 views
r/LocalLLaMA

Take Three: What’s the rub on memory sessions?

6/3/2026 · 50 views
r/LocalLLaMA

ui: Mermaid Diagrams in chat + interactive preview by allozaur · Pull Request #24032 · ggml-org/llama.cpp

6/3/2026 · 48 views
r/LocalLLaMA

Gemma 4 is coming - No Vision Tower - No Audio Tower

6/3/2026 · 47 views
r/LocalLLaMA

I developed a hard LLM Challenge

6/3/2026 · 42 views
r/LocalLLaMA

lipsync possible on mac?

6/3/2026 · 36 views
r/LocalLLaMA

Qwen 3.7 Plus just briefly appeared and then disappeared on OpenRouter.

6/3/2026 · 49 views
Reddit

Half the top 10 trending GitHub repos right now are "skills" projects, not models

6/3/2026 · 44 views
r/LocalLLaMA

Tensor split mode: CUDA error on latest llama.cpp with Qwen-3.6-27b

6/3/2026 · 34 views
r/LocalLLaMA

Calling it now Microsoft is buying Unsloth.

6/3/2026 · 42 views
r/LocalLLaMA

Helvete-nano

6/3/2026 · 31 views
r/LocalLLaMA

Holo3.1 35B/9B/4B/0.8B (Qwen 3.5 finetunes)

6/3/2026 · 37 views
r/LocalLLaMA

Mellum & Granite Embedding models are ready on llama.cpp

6/3/2026 · 38 views
r/LocalLLaMA

Another shout out to llama.cpp build b9455 2x3090

6/3/2026 · 42 views
r/LocalLLaMA

Microsoft Aion 1.0 Instruct and Aion 1.0 Plan models!

6/3/2026 · 39 views
r/LocalLLaMA

Nous Research — Hermes Desktop

6/3/2026 · 45 views
r/LocalLLaMA

Why do we benchmark quants on perplexity and prose but never on tool call validity?

6/3/2026 · 42 views
r/LocalLLaMA

Someone out there likely needs this

5/30/2026 · 34 views
r/LocalLLaMA

Everyone here self-hosts inference. Almost nobody self-hosts the tooling around it. That feels backwards to me.

5/30/2026 · 37 views
r/LocalLLaMA

Cost Analysis of my $6.4k Local LLM Server

5/30/2026 · 48 views
r/LocalLLaMA

Running Qwen 3.6 35b MoE With Zoo Code On M1 Max is Amazing! Fully local, battery-powered coding powerhouse!

5/30/2026 · 52 views
r/LocalLLaMA

Would a MacBook M5 16/24/32GB be an upgrade, complement, or waste next to my RTX 4060 laptop?

5/30/2026 · 38 views
r/LocalLLaMA

What features dramatically improved your custom memory system?

5/30/2026 · 38 views
r/LocalLLaMA

For those creating personal assistants locally - how has short/long term memory impacted your experience?

5/30/2026 · 46 views
r/LocalLLaMA

Parallax: Parameterized Local Linear Attention for Language Modeling

5/30/2026 · 47 views
r/LocalLLaMA

nvidia/Qwen3.6-35B-A3B-NVFP4 · Hugging Face

5/30/2026 · 39 views
r/LocalLLaMA

SupraLabs 50M Parameter Model Just Hit the Trending Page on Hugging Face 🤯

5/30/2026 · 38 views
r/LocalLLaMA

Why does Thinking Output More Tokens Than a Response?

5/30/2026 · 46 views
r/LocalLLaMA

[LLM analysis challenge] OPERATION: REVERSE ROBOTOMY. We need an LLM Neurosurgeon to extract a password from a fractured artificial mind.

5/30/2026 · 39 views
r/LocalLLaMA

Can't get over 250TPS on RTX5090 with Qwen3.5-4B

5/30/2026 · 39 views
r/LocalLLaMA

LFM2.5-8B-A1B release

5/30/2026 · 35 views
r/LocalLLaMA

anybody got llama-swap working answering concurrent requests for a single model?

5/30/2026 · 47 views
r/LocalLLaMA

STT -> LLM -> TTS pipeline

5/30/2026 · 44 views
r/LocalLLaMA

Qwen 3.6 coding choice–27B vs 35B quants

5/30/2026 · 33 views
r/LocalLLaMA

"What are you good at?"

5/30/2026 · 26 views
r/LocalLLaMA

Fulloch V2: 100% Local Voice Assistant for Home Assistant & Obsidian (Runs on 16GB VRAM)

5/30/2026 · 45 views
r/LocalLLaMA

MINISFORUM UM790 Pro

5/30/2026 · 35 views
r/LocalLLaMA

Gryphe/Pantheon-Reasoning-27B · Hugging Face

5/30/2026 · 34 views
r/LocalLLaMA

Open source : Turning vocal imitations into sound effects. (New UX for sound generation)

5/30/2026 · 43 views
r/LocalLLaMA

Vidai Community is now available: one Rust binary for cost attribution, guardrails and multi-provider routing on every LLM call

5/30/2026 · 39 views
r/LocalLLaMA

The best AI Model for Arabic dialects 🇪🇬🦅🧡

5/30/2026 · 40 views
r/LocalLLaMA

made a local voice AI for windows you can talk to in any language. open source, bring your own key

5/30/2026 · 41 views
r/LocalLLaMA

I have 2x PC's. One with a 5090 and one with a 4080. Is there an easy way to use both together networked?

5/30/2026 · 32 views
r/LocalLLaMA

Keeping multi-GPU rigs cool?

5/30/2026 · 32 views
r/LocalLLaMA

Breaking the music supply constraint

5/29/2026 · 34 views
r/LocalLLaMA

Uploaded my Qwen3.6 27B based fine tune, after two years of experience fine tuning models

5/29/2026 · 38 views
r/LocalLLaMA

Mutating Gemma 4 31B Dense in to a native Gemma 4 additive-MoE model

5/29/2026 · 45 views

How WeSearch handles this source

WeSearch's declared handling of r/LocalLLaMA's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.

Indexing: Allowed Snippet: Allowed AI summary: Limited Retrieval / RAG: Not asserted Model training: Not asserted Commercial reuse: Not permitted

More social sources

Visit r/LocalLLaMA directly →