Qwen/WebWorld 32B/14B/8B (Qwen3 finetune) (www.reddit.com)
model roundup
Qwen 3
-
WebWorld is a large-scale open-web world model series for training and evaluating web agents. It is trained on 1M+ real-world web interaction trajectories via a scalable hierarchical data pipeline, supporting: Long-horizon simulation (30+…
-
Wanted to see if a real voice loop — speak, model thinks, speaks back — could run entirely on a single device today, no cloud. Same codebase doubles as a live translator (speak in language A, hear it back in language B).
-
emotion-steering Extract and serve CAA-style emotion steering vectors for any HuggingFace causal LM, with a fast vLLM path for Qwen3. ┌────────────┐ ┌────────────────────┐ labeled │ extract │ vectors + AUC report │ serve │ contrasts ├─────…
-
A Qwen finetune, that feels VERY human (www.reddit.com)
Hello guys, So TL;DR, I was asked by multiple people to make an Assistant_Pepe_32B version, but the best base model contender was Qwen3-32B, a model that is very hard to tune on anything other than STEM. The concept of Assistant_Pepe is an…
-
dl.acm.org Performing security verification This website uses a security service to protect against malicious bots. This page is displayed while the website verifies you are not a bot.
-
Looking for Small VLM/MLLMs Alternatives to Qwen Series Models (www.reddit.com)
I have tried Qwen 3 VL family of models on my rtx3060, max I can load is Q8 8b. The task is visual reasoning/ instruction following.
-
Talk at Qwen Meetup Korea end of May. Looking for review on this draft before I build PPT slides off it.
-
Poor GPU Club : Tried Bonsai-8B on CPU & CUDA (www.reddit.com)
Got a chance to check this model today. 8GB VRAM(RTX 4060 Laptop GPU) & 32GB DDR5 RAM.
-
Hi everyone, wanted to see how far QVAC could be pushed on a phone: full speech-to-text → LLM → text to-speech running locally, no network, and get it close to a real conversation. Stack (Android, all via qvac sdk): - STT: Parakeet (stream…
-
https://preview.redd.it/7yei65sbugyg1.png?width=1703&format=png&auto=webp&s=ad388c51dd10cb44b41a99876d28797e006fd138 Stanford's Generative Agents = one LLM cosplaying 25 personas. I wanted agents that actually become different people — dif…
-
Best RTX Pro 6000 vllm settings? (www.reddit.com)
Just got myself (for my company) a RTX Pro 6000 Blackwell Workstation card. Managed to get really good TPS on qwen3 27b fp8.
-
Introducing Chirp (www.reddit.com)
Hey everyone, I’ve been working on Chirp, a native offline text-to-speech desktop app. It runs locally on your machine, supports both Kokoro and Qwen3-TTS, and is written in C++ and Rust.
-
Why aren't people using omni models for speech agents? (www.reddit.com)
I've been benchmarking open source omni models like Qwen3-Omni for speech to speech tasks and they perform... really well.
-
Show HN: I read Replika's privacy policy and then built a competitor (apps.apple.com via hn)
I'm genuinely surprised at what people are willing to share with AI companions. Read Replika's privacy policy.
-
How do you actually use Qwen3 72B Instruct locally? (www.reddit.com)
I just got Qwen3 72B Instruct running on a high RAM setup and I’m kinda confused about the proper way to use it. What’s the correct workflow for running it smoothly (like best quant, tools, or runtime)?
-
Going through university right now, and we have massive 100 page pdfs/ppts with soo much fluff that its annoying to go through. until now ive been using chatgpt for it, but realized that the output tokens are HEAVILY limited, and loses a L…
-
If you have low vram - qwen 3 tts is good If you need something unique go for - tada 3b but it need 28gb vram If you want best tts rn + have the commercial use allowed then go for - moss tts 8b its literally the best model out there Litera…
-
Heya guys and gals, Around a year ago I released and posted about Persona Engine as a fun side project, trying to get the whole ASR -> LLM -> TTS pipeline going fully locally while having a realtime avatar that is lip-synced (think VTuber)…
-
LLM speed t/s (www.reddit.com)
All I see is "it gives me **/s bla bla bla" all together with q4, q3... even when chatting with qwen3.