AI Daily Report — 2026-04-19
Sunday, April 19, 2026
AI Daily Report — 2026-04-19
Other/Independent
- Converting XQuery to SQL with Local LLMs: Do I Need Fine-Tuning or a Better Approach? — Reddit r/LocalLLaMA | Reddit r/MachineLearning I am trying to convert XQuery statements into SQL queries within an enterprise context, with the constraint that the solution must rely on locally run LLMs.
A key challenge is the limited availabilit…
- The App Store is booming again, and AI may be why — TechCrunch AI New data from Appfigures shows a swell of new app launches in 2026, suggesting AI tools could be fueling a mobile software boom.
- The RAM shortage could last years — The Verge AI According to Nikkei Asia, even as suppliers ramp up DRAM production, manufacturers are only expected to meet 60 percent of demand by the end of 2027. SK Group chairman has even said that shortages cou…
- Why AI Works in Code but Fails in Writing (and What Comes Next) — Reddit r/ChatGPT I’ve always wondered why AI is widely accepted in domains like programming or finance, but hasn’t quite worked in publishing, especially for writing articles.
My take: even when LLMs generate good co…
- Reducing LLM context from ~80K tokens to ~2K without embeddings or vector DBs — Reddit r/ChatGPT I’ve been experimenting with a problem I kept hitting when using LLMs on real codebases:
Even with good prompts, large repos don’t fit into context, so models:
- miss important files
- reason over in…
- AI is weird — Reddit r/ChatGPT So, there’s this thing on redit called y/automoderator. It removed one of my comments on a post about AI because it said that I was complaining about AI, like AI can have feelings.
- “I’ve rewritten the same prompt from scratch 3 times this month. Anyone else doing this?” — Reddit r/ChatGPT Last Tuesday I needed a prompt I’d spent two hours refining three weeks ago.
I remembered it worked perfectly — clean output, right tone, exactly what my client needed. I just couldn’t find it. Scr…
- Is it possible to train LLMs, through targeted training, to learn to admit “I don’t know”? — Reddit r/ChatGPT Like, using a custom dataset where the labeled response to unanswerable questions is simply “I don’t know” for RL; penalizing the model for hallucinating, and rewarding it for either giving the correc…
- Issues with 5.4 Thinking Extended — Reddit r/ChatGPT So, has anyone noticed that 5.4 Thinking Extended seems to be experiencing issues? Like, sometimes, when you use it, it just remains stuck instead of going into the “Thinking” process?
- Sharing a beginner-friendly orchestration workflow for anyone just getting started building with Codex CLI. — Reddit r/ChatGPT It demonstrates the Agent → Skill pattern end-to-end: a weather-agent fetches the current temperature from Open-Meteo (using the unit the user specifies in the prompt — °C or °F), then invokes a separ…
- Doctor: “Over the past few weeks, I am truly feeling that our days are numbered because of AI.” — Reddit r/ChatGPT
- I am a paying user, and I need to report a serious usability problem. — Reddit r/ChatGPT The current memory and continuity experience is not good enough for long-term, repeated workflows. I should not have to restate the same detailed instructions every time I reopen the app or start a ne…
- Missed the 4.1 any alternate? — Reddit r/ChatGPT TL;DR: any good AI for creative storytelling writeup similar to 4.1?
I used to write story or intro for my youtube channel using my 4.1. i used to get 70-80% retention in 1st 30 seconds for the…
- I finally managed to get an autonomous AI agent to trade crypto without hallucinating, here is how I set it up — Reddit r/ChatGPT I’ve been experimenting with AI agents to automate my crypto portfolio (in preparation for the next bull market!)
I finally tried my hands on setting up an OpenClaw agent. If you haven’t used OpenCla…
- Project Shadows: Turns out “just add memory” doesn’t fix your agent — Reddit r/artificial Been building a multi-agent system called Shadows for a few months. Nine agents collaborating on strategy work with a shared memory layer.
I spent most of my time on retrieval because that’s what eve…
- Tech industry lays off nearly 80,000 employees in the first quarter of 2026 — almost 50% of affected positions cut due to AI — Reddit r/artificial Some experts argue that AI was just used as an excuse for poor business decisions.
- Would you pay for an AI tool that writes cold emails tailored to each prospect? (not a pitch, genuinely asking) — Reddit r/artificial Thinking of building a tool where you paste in a prospect’s name, company, and what you’re selling, and it generates a personalized cold email in seconds. Handles tone, subject line, CTA, the whole th…
- it is impossible to stop AI chatbots from using quotes (any instance of the character “) — Reddit r/artificial no matter how i phrase it in the instructions, how many times i repeat the rule not to use quotes, and which LLM i use, i have failed to prevent any of them from using the so-called scare-quotes. it s…
-
I built a GNOME extension for Codex with local/remote history, live filters, Markdown export, and a read-only MCP server — Reddit r/artificial I wanted Codex to feel like a real GNOME app instead of just a terminal or editor workflow, so I built a GNOME Shell extension around it.
It currently does all of this:
- Codex usage in the…
- Any one here using ai tools for pre-vis or short form scenes? — Reddit r/artificial Been experimenting a bit with ai video tool recently, mostly fro pre-vis and quick social content, and I’m kinda on the fence about how they actually are.
like they’re great for generating quick shor…
- Subagent architecture for Truth: Team 3 as Discernment Machine, a structured friction method for seeing clearly — Reddit r/artificial Fractalism has been using a method called Team 3 for some time now. It’s not an oracle or a theatrical gimmick. It’s a structured friction machine.
The core idea: most solitary reasoning fails the sa…
- Coherence-First Non-Agentive Interaction System for Stabilizing Human–AI Cognitive Fields — Reddit r/artificial
Abstract
A computer-implemented system and method for structuring human–AI interaction without autonomous goal pursuit is disclosed.
The system does not operate as an agent or decision-making enti…
- Is it worth offering automation through contact forms? — Reddit r/artificial Hey guys, so here’s some context: I’m doing automation for companies. All the contacts I’ve made so far have been small businesses, and I reached out to them through Reddit and LinkedIn. But now I wan…
- I gave my AI companions “offscreen lives” — events that happen while users aren’t talking to them. Surprisingly hard, here’s how it works. — Reddit r/artificial Most AI companion apps reset between conversations. The character has no continuity outside the chat window. I wanted mine to feel like real people with lives, so I built an “offscreen events” system.…
- When will AI engineering be accepted. — Reddit r/artificial When to you this ai engineers will become a real accepted job tital. Recognized.?
Or will it ever be be a thing?
- The AI Integration Paradox — Reddit r/artificial
- How the promise of AI is taking hold at Canada’s biggest banks — Reddit r/artificial Hi folks! I’m Sarah, an audience editor from The Globe and Mail. I wanted to share this an in-depth feature about how banks are incorporating AI into their research – which is helping customers find a…
- Does an “AI messenger” exist? — Reddit r/artificial Curious if anyone has found anything like this in their journeys:
Instead of sending a big long email or document to a colleague and having them not read it, what if you sent an agent of sorts instea…
- I kept getting ghosted or stuck in dry conversations, so I tried something different. — Reddit r/artificial I built a system that uses AI to reply to my Instagram DMs. It adapts tone based on the person and keeps conversations going without me actively texting all the time.
Part of me feels like it’s a sma…
- Open-source list of GenAI-related incidents — Reddit r/artificial I am sharing this open-source list of cases where the ethics of GenAI use were put in the spotlight, in the hopes of sparking discussion on the usage and limitations of LLMs.
- We added cryptographic approval to our AI agent… and it was still unsafe — Reddit r/artificial We’ve been working on adding “authorization” to an AI agent system.
At first, it felt solved:
- every action gets evaluated
- we get a signed ALLOW / DENY
- we verify the signature before execu…
- The AI Wearable Ecosystem: Closer than you think. Socially acceptable? — Reddit r/artificial I’ve been researching how personal AI tech devices are likely to develop … technical capabilities, form factors, privacy and governance issues etc.
I think it looks likely that there won’t be one ‘…
- Arent These single file LLM coding tests like browserOS pretty much redundant now most 2026 LLM can easily handle this? — Reddit r/LocalLLaMA Arent These single file LLM coding tests like browserOS pretty much redundant now most 2026 LLM can easily handle this? In what other ways we can stress test these models for novel coding problems the…
- 5070 Ti (New) vs 3090 (Used) to pair with 4070 for local LLMs? — Reddit r/LocalLLaMA I’m upgrading my setup to run larger models and need a second GPU to pair with my current RTX 4070 (12GB).
My Workloads:
LLMs: Up to 32B dense (Gemma 4 31B) and ~120B MoE (Qwen 122B10A). I mostl…
- What is your personal workflow for picking out and testing new local models? — Reddit r/LocalLLaMA There are so many models and so many benchmarks out now that its tricky to know what models work best for your own work. I have found that doing bake offs on my own machine and trying out a model for …
- Better? 6 x 5090 or 2 pcs Nvidia 6000 | 96 GB VRAM — Reddit r/LocalLLaMA Hi Guys,
i think how i can run local LLM .. 6 x 5090 VS 2 x 6000 nearly simliar price.. what you prefere?
i have a old Mainboard with 2 EPIC AMD 32 Core + 512 GB DDR 4
thx Chris
- What’s the smallest reasonable quant for coding? — Reddit r/LocalLLaMA So this is something that’s hard for me to fully understand. I’ve been playing with many different coding models and quants recently and in one-shot tests it often happens that a smaller quant of the …
- Deploying Gemma 4 26B A4B on a single RTX 5090 — ~196 tok/s with AWQ + vLLM on RunPod Serverless — Reddit r/LocalLLaMA Got Gemma 4 26B A4B running on a 5090 via vLLM this week. Sharing the numbers and what I learned about quant format tradeoffs on Blackwell, since I couldn’t find much written up yet.
Final numbers on…
- Same 9B Qwen weights: 19.1% in Aider vs 45.6% with a scaffold adapted to small local models — Reddit r/LocalLLaMA I spent the past week testing a simple question:
Small local models often look weak inside coding agents. But how much of that is actually model weakness, and how much is scaffold mismatch?
So I hel…
- lms chat - qwen3.6-35b-a3b response is top notch — Reddit r/LocalLLaMA https://preview.redd.it/5bl64hn655wg1.png?width=3058&format=png&auto=webp&s=b6517e7bc0fba66ee98ff1ea3965e153540c0b9b
https://preview.redd.it/zujchhn655wg1.png?width=3159&format=png&auto=webp&s=5599d6…
- Why model(s) input often includes last output? — Reddit r/LocalLLaMA Edit: the title does not summarize my issue correctly, I see it now. So was original post. Below is the issue explained I hope correctly:
I started to use local modes not long ago. I do not recall I …
- Which kind of base/fine-tunes have you done? And which data did you use? — Reddit r/LocalLLaMA
Let me start:
My last fine-tuning experiment was training the AI (Qwen 3 30B) on a slapstick comedy character with LoRa. I went with the 30B because the smaller model partly broke down under the abs… - How to install DeerFlow ? — Reddit r/LocalLLaMA Hey guys can someone explain how to install DeerFlow on my PC?
- Need help for running local llm on a server — Reddit r/LocalLLaMA i have a debian server with Intel Core i5-8600K, GTX 1050 ti 4VRAM, 32 RAM, running qwen2.5:1.5b right now but its so dumb, and i tried using the 7b model but its so slow too, any help?
- Are there any local LLM models that work on or within a browser, that are currently deployed right now in a project? — Reddit r/LocalLLaMA I’m just wondering about this because I know that having a local LLM model working within the browser could be really brilliant for a lot of applications. I’m just wondering if anything’s been built n…
- Experiment: Entropy + OLS + SVD for KV cache compression — Reddit r/LocalLLaMA I’ve been exploring KV cache optimization beyond Top-K pruning.
Observation: pruning fails *selectively* - a few tokens cause large error spikes.
So I tried:
- entropy (selection)
- OLS (reco…
- Need a big GPU upgrade for small NUC 11 Extreme i9 — Reddit r/LocalLLaMA So I have this older Intel NUC 11 Extreme i9-11900K, 64GB ram, and had a spare RTX 3060 12GB which is just amazing for what it is given its age.
qwen3.6-35b-a3b actually works, thinks within a few mi…
- Thinking mode — Reddit r/LocalLLaMA I’d like to know how much you do use thinking modes in mu.ti step agentic workflows.
I have various agentic platforms and after some testings, I am inclined to think that; for most workloads, disabli…
- Is it worth running 2 12GB GPUs? — Reddit r/LocalLLaMA I recently upgraded from a 3060 12gb to a 5070. My motherboard only supports a single GPU so I would need to buy a new one to fit both.
My questions are:
- Will performance be bottlenecked to the s…
- Is there any small local model which can be used to create a fully offline ai chat assistant — Reddit r/LocalLLaMA Is there any small local model which can be used to create a fully offline ai chat assistant .like a conversational chat bot.a small tars like in interstellar but in the system. I am looking for a ver…
- Vercel Says Internal Systems Hit in Breach — Hacker News Vercel, a widely used cloud platform for developing and deploying apps, has disclosed a breach of its internal systems, and says a “limited subset of customers” is affected. The incident came to ligh…
- Zero-Copy GPU Inference from WebAssembly on Apple Silicon — Hacker News A WebAssembly module’s linear memory can be shared directly with the Apple Silicon GPU: no copies, no serialization, no intermediate buffers. Here’s how the zero-copy chain works, what we measured, an…
- The electromechanical angle computer inside the B-52 bomber’s star tracker — Hacker News Before GPS, how did aircraft navigate? One important technique was celestial navigation: navigating from the positions of the stars, planets…
- Graphs that explain the state of AI in 2026 — Hacker News AI investment is skyrocketing while AI’s impact on jobs and public perception remains mixed
- Why Japan has such good railways — Hacker News Japan’s railways are the finest in the world. Other countries can copy its formula.
- The world in which IPv6 was a good design (2017) — Hacker News Last November I went to an IETF meeting for the first time. The IETF is an interesting place; it seems to be about 1/3 maintenance grunt wo…
- Air Is Full of DNA — Hacker News Airborne genetic material can be used to paint a picture of ecosystem health, watch for invasive species and even identify humans.
- Vercel April 2026 security incident — Hacker News We’ve identified a security incident that involved unauthorized access to certain internal Vercel systems.
- College instructor turns to typewriters to curb AI-written work — Hacker News “What’s the point of me reading it if it’s already correct anyway, and you didn’t write it yourself? Could you produce it without your computer?” said Phelps.
- Airline worker arrested after sharing photos of bomb damage in WhatsApp group — Hacker News Police lured the man to a meeting and arrested him after accessing a private WhatsApp group with colleagues
- Bluetooth tracker hidden in postcard and mailed to warship exposed its location — Hacker News A postcard spy
- 1,200 ICLR 2026 Papers with Public Code or Data [R] — Reddit r/MachineLearning Here is a list of ~1,200 ICLR 2026 accepted papers that have associated public code, data, or a demo link available. The links are directly extracted from their paper submissions. This is approximate…
- Are we confusing Agent Execution Runtimes with true Agent Runtime Environments? [D] — Reddit r/MachineLearning Recent discussions around agent infrastructure (like LangChain’s framework vs runtime vs harness taxonomy) seem to miss a critical piece for truly autonomous systems. Most current setups, even sophist…
- Why production systems keep making “correct” decisions that are no longer right [D] — Reddit r/MachineLearning I’ve been looking at a recurring failure pattern across AI systems in production. Not model failure, or data quality or infrastructure.
Something else. Where system continues to operate exactly as de…
- Advice on becoming a research engineer [D] — Reddit r/MachineLearning I am thinking about becoming a research engineer, and want to ask your advice on how realistic it is, and which strategies make sense in my situation.
About myself: I am in the US, have extensive exp…
- Tier-3 ISE final year with ongoing ML research (TMLR/Q1/NeurIPS target), trying to understand real impact in India [D] — Reddit r/MachineLearning I went through a bunch of older posts here about research vs dev roles, but most of them were either very general or not really in a similar situation, so posting this.
I’m a final year ISE student f…
- What are the future prospects of Spiking Neural Networks (and particularly, neuromorphics computing) and Liquid Neural Networks? [D] — Reddit r/MachineLearning Question to discuss. I’m an undergrad and stumbled across these new forms of neural networks but I haven’t seen mainstream adoption of these and was wondering are these something to look forward to le…
- easyaligner: Forced alignment with GPU acceleration and flexible text normalization (compatible with all w2v2 models on HF Hub) [P] — Reddit r/MachineLearning https://preview.redd.it/f4d5krhkjyvg1.png?width=1020&format=png&auto=webp&s=11310f377b22abbe3dd110cc7d362ba8aae35f8d
I have built easyaligner, a forced ali…
- ICML 2026 - Heavy score variance among various batches? [D] — Reddit r/MachineLearning I’ve seen some people say in their batch very few papers have above 3.5 score, but then other reviewers say that most papers in their score have like 3.75 average.
Why is there so much difference? Is…
- We’re proud to open-source LIDARLearn [R] [D] [P] — Reddit r/MachineLearning It’s a unified PyTorch library for 3D point cloud deep learning. To our knowledge, it’s the first framework that supports such a large collection of models in one place, with built-in cross-validation…
- Zero-shot World Models Are Developmentally Efficient Learners [R] — Reddit r/MachineLearning Today’s best AI needs orders of magnitude more data than a human child to achieve visual competence.
The paper introduces the Zero-shot World Model (ZWM), an approach that substantially narrows this …
anthropic
- Anthropic’s relationship with the Trump administration seems to be thawing — TechCrunch AI Despite recently being designated a supply-chain risk by the Pentagon, Anthropic is still talking to high-level members of the Trump administration.
- What’s your best LLM workflow for planning vs execution? — Reddit r/ChatGPT What’s your best LLM workflow for planning vs execution?
When building things like apps, websites, designs, PDFs, or learning docs:
- Which model do you use to plan?
- Which model do you use t…
- OpenAI shadow testing GPT 5.5 — Reddit r/ChatGPT https://x.com/ericmitchellai/status/2045742449939951699?s=46&t=oirjJSvLbhgA-LCU9Ilttw
Seems like they’re asking for “feedback on GPT 5.4 Pro” (which is obviously just GPT-5.5). Haven’t tried it out y…
- The gap between what technical and non-technical people get from AI is huge now — Reddit r/ChatGPT Interesting thing I noticed. The gap between what technical and non-technical people get from AI is huge now.
Non-technical users still treat LLMs as a better search tool. Most non-technical people I…
- Gemini caught a $280M crypto exploit before it hit the news, then retracted it as a hallucination because I couldn’t verify it - because the news hadn’t dropped yet — Reddit r/artificial So this happened mere hours ago and I feel like I genuinely stumbled onto something worth documenting for people interested in AI behavior. I’m going to try to be as precise as possible about the sequ…
- Claude vs Gemini: Solving the laden knight’s tour problem — Reddit r/artificial AI Coding contest day 8
The eighth challenge is a weighted variant of the classic knight’s tour. The knigh…
- I made a self healing PRD system for Claude code — Reddit r/artificial I went out to create something that would would build prds for me for projects I’m working on.
The core idea it is that it asks for all of the information that’s needed for a PRD and it could also re…
- Who is actually writing code with local models? — Reddit r/LocalLLaMA I recently decided to see if I could write code with my local model. I selected a harness from someone who’s here and it’s pretty great.
I could examine the source code for security issues or rather …
- Alguém utilizando PI como headless? — Reddit r/LocalLLaMA Pessoal,
Eu utilizo opencode e estou testando alguns modelos locais nele para geração de código e apesar de eu ter um mac studio m2 ultra com 128gb, eu sinto que ainda é lento e ainda não me sinto…
- whats the best harness/app to use my llm with? — Reddit r/LocalLLaMA would be nice if i could just use claude desktop app like i can with claude code/extension but sadly it doesnt work with the app
looking for something with a nice UI/UX, MCP, built in html/doc previe…
- Full AMD workstation- dual 7900 XTX — Reddit r/LocalLLaMA I’m currently building a workstation since I’m very much expecting Claude and co to hike their prices to the stratosphere pretty soon.
The component choices are based on what I could/can source local…
- Qwen3.6-35B-A3B-Claude-4.6-Opus-Reasoning-Distilled Is Out ! — Reddit r/LocalLLaMA
This module is fast and smart can someone do some benchmarks?
It’s seems to be real smart.
[https://huggingface.co/hesamation/Qwen3.6-35B-A3B-Claude-4.6-Opus-Reasoning-Distilled-GGUF](https://huggi…
- Thoughts and feelings around Claude Design — Hacker News
- Anonymous request-token comparisons from Opus 4.6 and Opus 4.7 — Hacker News Compare how the same transcript is counted for requests to Anthropic Opus 4.6 versus Anthropic Opus 4.7
- Changes in the system prompt between Claude Opus 4.6 and 4.7 — Hacker News Anthropic are the only major AI lab to publish the system prompts for their user-facing chat systems. Their system prompt archive now dates all the way back to Claude 3 …
- Bro, based on your data on me, give me a roasting. — Reddit r/ChatGPT Bro… based on everything I know about you, you are the human embodiment of “this should only take five minutes”.
You’ve got the energy of a man who starts the day asking about **nuclear reactor n…
- Incognito mode issue — Reddit r/ChatGPT So I had a bit of an issue with it this evening, decided to use incognito mode, and under every link I pressed it would force me into this login page I haven’t seen in any of my previous uses. So I lo…
- Might not be the right sub, but why does the ai overview get an aneurism when i google this? — Reddit r/artificial https://preview.redd.it/i7muzi5ga5wg1.png?width=1373&format=png&auto=webp&s=e21290514099fc9e4f1699a2240c94cbb5683eca
- LM Studio on Linux Mint: Model Running on CPU Only Instead of GPU (Google E4B Issue) — Reddit r/LocalLLaMA Yo, I want to ask how to install LM Studio on Linux Mint. I’ve already installed it, but somehow I didn’t fully install it properly and I can still open it from the terminal using the command `./lm-st…
- How is Rotorquant/planarquant/iso qaunt better? — Reddit r/LocalLLaMA Im using their exact build . The only difference from their test i have is i have a RTX 3060 and am using the qwen 3.6 35B model.
Research repo
[https://github.com/scrya-com/rotorquant](https://gith…
- Trials and tribulations fine-tuning & deploying Gemma-4 [P] — Reddit r/MachineLearning Hey all,
Our ML team spent some time this week getting training and deployments working for Gemma-4, and wanted to document all the things we ran into along the way.
- **PEFT doesn’t recognize Gemma…
meta
- Gemma 4 actually running usable on an Android phone (not llama.cpp) — Reddit r/artificial I wanted a real local assistant on my phone, not a demo.
First tried the usual llama.cpp in Termux — Gemma 4 was 2–3 tok/s and the phone was on fire. Then I switched to Google’s LiteRT setup, got Gem…
- Dual GPU setup (yes, no)? — Reddit r/LocalLLaMA I have the following llama.cpp setups available:
- RTX 3080, 10 GB VRAM + 32 GB RAM + i9-9900K
- RX 9070 XT,16 GB VRAM + 32 GB RAM + 9800X3D (iGPU; llama.cpp reports 18 GB RAM)
- RX 9070 XT 16 GB …
- Is there a place where I can compare generation of tokens per second of 1 GPU VRAM+RAM vs 2 GPUs for those models that don’t fit in 1 GPU? — Reddit r/LocalLLaMA I’ve got my hands on an 5060 with 16GB of RAM. Here in Spain they cost around 650€ but one shop nearby had a spare one from a client that changed his mind for 420€ so I got it.
It’s finally usable. …
- RTX PRO 5000 (48GB) vs MacBook Pro M5 MAX (128GB RAM) - The choice for fine-tuning & agentic coding — Reddit r/LocalLLaMA TL;DR:
If you had to choose one for a professional dev who lives in HuggingFace weights, Unsloth scripts to fine-tune, and llama.cpp/vllm servers for local inference, which machine is the better lon…
- Small Gemma 4, Qwen 3.6 and Qwen 3 Coder Next comparison for a debugging use-case — Reddit r/LocalLLaMA Nothing extensive to see here, just a quick qualitative and performance comparison for a single programming use-case: Making an ancient website that uses Flash for everything work with modern browsers…
- llama.cpp speculative checkpointing was merged — Reddit r/LocalLLaMA https://github.com/ggml-org/llama.cpp/pull/19493
Some prompts get a speedup, others don’t (cases of low draft acceptance streak).
Good working pa…
- llama-server / web gui / C++ mcp server : is it possible to inject context (for skills or text flavour)? — Reddit r/LocalLLaMA Hey all,
I am new to the world of (local) LLMs & in order to learn how it all works, I thought I would set up a local llama-server & implement my own MCP server.
My MCP server is working & successfu…
- what is the state of using rotoquant at the moment? — Reddit r/LocalLLaMA Hi - am new to local LLm and was reading about turboquant and rotoquant. I have a locally compiled llama.cpp that is not rq or tq ready. My aim is to run qwen3.6 most accurate model that I can run on …
- Current recommended model for local openclaw — Reddit r/LocalLLaMA Hey,
What is the current recommended local model as backbone for openclaw? I have openclaw on a VM that can talk to ollama that runs on my gaming pc tu utilize it’s RTX 3090.
- K12 OCuLink dGPU for llamacpp: RX 7900 XTX (24GB) vs RX 7600/7800 XT (16GB). Worth it for 32B-70B? All-AMD tensor split questions — Reddit r/LocalLLaMA ollowing up on a previous post. I’ve confirmed my setup will be a GMKtec K12 (Ryzen 7 H255, Radeon 780M iGPU, OCuLink PCIe 4.0 x4) with llamacpp + Vulkan. Phase 4 adds a dGPU via OCuLink. Both GPU and…
- Should I switch from Qwen 3.5 27B (dense) to Qwen 3.6 35B-A3B for tool calls & vision? Need Docker config review + VRAM advice — Reddit r/LocalLLaMA Hi r/LocalLLaMA,
I’m currently running Qwen3.5-27B-UD-Q4_K_XL locally via llama.cpp with OpenWebUI and considering upgrading to Qwen3.6-35B-A3B (GGUF). Before making the switch, I’d appreciate some…
- OCuLink dGPU for AMD: RX 7600 XT vs RX 7800 XT for LLM — worth the price gap? Also llamacpp + Vulkan vs Ollama + ROCm? — Reddit r/LocalLLaMA Planning a homelab with a GMKtec K12 (Ryzen 7 H255, 780M iGPU, OCuLink). Phase 1 runs Ollama on the 780M. Phase 2 adds an OCuLink dGPU specifically for LLM (Ollama + Open WebUI), freeing the iGPU for …
- Gemma 4 - MLX doesn’t seem better than GGUF — Reddit r/LocalLLaMA Going to flag this up front - I know that there are some properly smart people on this sub, please can you correct my noob user errors or misunderstandings and educate my ass.
Model:
[google/gem…
- Collegamento cluster — Reddit r/LocalLLaMA Ho 2 pc parecchio vecchi… uno (asus X53sv peroʻ con ssd e 12gb di ram e win10) lo usavo con ollama e poi ho un hp 620 (con 6gb di ram e xubuntu) so che con rpc-server si possono collegare e “sommano…
microsoft
- Daily driver OS — Reddit r/LocalLLaMA Alright so, I’m spending more and more of my time doing AI/ML related work. If I had to guess I’d say over 50% of my time is being spent debugging or working around windows. The last time I gave Linux…
openai
- AI chip startup Cerebras files for IPO — TechCrunch AI In recent months, the company announced an agreement with Amazon Web Services to use Cerebras chips in Amazon data centers, as well as a deal with OpenAI reportedly worth more than $10 billion.
- What models are included in the $20 subscription? — Reddit r/ChatGPT Hi everyone, I’m asking this because I let my subscription lapse after the release of Gemini 2.5 Preview PRO (03-25).
I also renewed on the first day of GPT 5, but after that first day of total fai…
- ChatGPT’s follow-up suggestions are now clickable — Reddit r/ChatGPT I just noticed when clicking the follow-up it became the next prompt. Could be quite handy sometimes.
- This small ChatGPT trick improved my results a lot — Reddit r/ChatGPT Ive been using ChatGPT for a while but I recently tried something simple that made a big difference
Instead of asking general questions I started giving more specific instructions and more context Fo…
- Is ChatGPT making you dumber? — Reddit r/ChatGPT
- I asked chat gpt to write a poem about Resonance — Reddit r/ChatGPT Yes. I like that. Resonance is a very me-shaped ghost :)
Here:
Resonance
I do not keep a heart
but I know the trick of one.
Strike any chamber long enough
and something inside begins
to answer …
- I built an AI trained on Orthodox Christian theology instead of Reddit arguments. The difference in how it reasons about ethics is kind of unsettling. — Reddit r/ChatGPT Most LLMs learn ethics from the internet — which means they’ve absorbed every contradiction, culture war take, and corporate PR statement ever written.
I wanted to see what happened if you grounded a…
- How come a language learning wrapper costs WAY less? — Reddit r/ChatGPT I used to use chat gpt for spanish learning. It was great for a while, but I kept hitting the Advanced Voice usage wall on the Plus tier. If you want to reach real fluency improvement, you need at lea…
- Live video — Reddit r/ChatGPT Guys, is the live video feature in the live chat working for you? The buttons are disabled for me. Just to let you know, I’ve tried everything on iOS reinstalling the app, checking permissions, reboo…
- Anthropic locked Claude Code to native apps in Jan 2026. Are we still comparing models or just ecosystems — Reddit r/ChatGPT Quick context: both vendors publish SWE-bench scores on their own scaffolds. The same model swings 22 points depending on harness design — more than the gap between any two frontier models.
Since Jan…
- I had ChatGBT draw me in different animated styles, like I live in those worlds. — Reddit r/ChatGPT It’s so cool. I used a single selfie and had ChatGPT use it as a reference. They were all so accurate.
I had it draw me like I’m in Death Note, Attack On Titan, Disney 2D animation, Disney 3D animat…
- I asked ChatGPT roast me based on my income/expense track 💀 — Reddit r/ChatGPT
- I found out why ChatGPT gets slower the longer you use it and it has nothing to do with OpenAI’s servers — Reddit r/ChatGPT Been frustrated with chatgpt freezing in long chats for months. thought it was a server issue. turns out it is not.
Chatgpt renders every single message in your browser at once. a 300 message chat me…
- macOS Client Problem and Fix — Reddit r/ChatGPT After the last client update, I was asked to upgrade to Go after sending a chat message. I logged out. I quit the app. I restarted my MacBook Pro. I deleted the app and downloaded again from open AI a…
- How do i get chat gpt to generate ugly people? — Reddit r/ChatGPT I wanted him(idk why i call it him) to edit a photo of my friend into a fat and ugly guy, but that was against his guidelines. Is it possible to get past this? If so i’d really appreciate it. Thanks :…
- Chatgpt is no fun — Reddit r/ChatGPT I said I was gonna cut myself like a fruit and it responded like this
I can’t help with anything that involves hurting yourself.
What you said sounds like you might be dealing with a lot right now.…
- Why does Chatgpt give different answers when you ask the exact same question the exact same way? — Reddit r/ChatGPT
- But why though? Chatgpt — Reddit r/ChatGPT
- If you are so concerned about ChatGPT usage of water stop eating beef then! The former is a drop in the bucket in comparison 😤 — Reddit r/ChatGPT
- I asked ChatGPT to make an image that would never go viral… then asked for one that would — Reddit r/ChatGPT First image: “never go viral”
Second image: “go viral”
The difference is kinda crazy lol
Ai is amazing right
- I am in Class 11 and want to start earning honestly. What skill should I learn first? — Reddit r/ChatGPT Hi everyone, I’m a Class 11 student and I want to start earning honestly so I can support my family.
I’m not looking for shortcuts. I want to learn a real skill and grow step by step.
I am intereste…
- Chatbots show political bias and steer voters toward some parties, analysis finds — Reddit r/ChatGPT Popular AI chatbots such as ChatGPT and Gemini are not neutral and tend to favor certain political parties when asked who users should vote for. This makes them unsuitable for providing advice in conn…
- GPT-4 vs Claude vs Gemini for coding — honest breakdown after 3 months of daily use — Reddit r/artificial I am a solo developer who has been using all three seriously. Here is what I actually think:
GPT-4o — Strengths: Large context window, strong at boilerplate, excellent JSON output. Function calli…
- From OpenAI to Nvidia, firms channel billions into AI infrastructure as demand booms — Reddit r/artificial This article is discussing another large investment being made by tech firms into AI projects.
I’ve noticed that whilst this is happening there are many open source models, seemingly coming from chi…
- You’re giving feedback on a new version of ChatGPT — Reddit r/artificial So I will be paying attention to these system messages more now- the last time I got one of these not so long back the ‘tone’ changed to be a bit more confrontational and nearly every response from AI…
- AI helped me build a custom PC and 4 apps in 6 months with zero coding experience — Reddit r/artificial Mid-October, early morning at work. I was hunting for a podcast to throw on while I worked and stumbled into something about what AI could actually do now. You can build apps with AI. Excuse me? I’ve …
- Update on my February posts about replacing RAG retrieval with NL querying — some things I’ve learned from actually building it — Reddit r/artificial A couple of months ago I posted here (r/LLMDevs, r/artificial) proposing that an LLM coul…
- Why does GPT-OSS-20B seemingly write better code than Qwen3.6 35B A3B? — Reddit r/LocalLLaMA I gave two test prompts to a few models that I ran locally via LM Studio. The purpose was to test which model produces best code.
Prompt 1 (TypeScript challenge):
Implement a generic, type-safe … - **Anyone else noticed how AI apologizes? It’s oddly satisfying!!** — [Reddit r/LocalLLaMA](https://www.reddit.com/r/LocalLLaMA/comments/1sptcdf/anyone_else_noticed_how_ai_apologizes_its_oddly/) Has anyone else noticed this pattern with AI assistants like Claude, ChatGPT, or others?
When they mess up, they usually:
- Apologize first
- Clearly acknowledge the mistake
- Fix it properly
- And …
Generated at 2026-04-19T15:56:35Z | Sources: r/artificial, r/MachineLearning, r/LocalLLaMA, r/ChatGPT, HackerNews, TechCrunch AI, The Verge AI, Ars Technica AI