AI Daily Report — 2026-05-04
Monday, May 04, 2026
AI Daily Report — 2026-05-04
Other/Independent
- ‘This is fine’ creator says AI startup stole his art — TechCrunch AI The ad comes from Artisan, the AI startup behind billboards urging businesses to “stop hiring humans.”
- In Harvard study, AI offered more accurate emergency room diagnoses than two human doctors — TechCrunch AI A new study examines how large language models perform in a variety of medical contexts, including real emergency room cases — where at least one model seemed to be more accurate than human doctors.
- How the internet’s favorite squirrel dad made the hottest camera app of 2026 — The Verge AI It’s not hyperbole to call DualShot Recorder an overnight sensation. It took only 12 hours from the time it was released to hit number one on the App Store’s list of top paid apps. It was a surprise s…
- AI music is flooding streaming services — but who wants it? — The Verge AI This is The Stepback, a weekly newsletter breaking down one essential story from the tech world. For more on how AI is changing music and the music industry, follow Terrence O’Brien. The Stepback arri…
- AI told me one patient had 148 pregnancies — Reddit r/artificial I was working with medical dataset (diabetes UCI data) and I was using AI data analyst . I asked AI to load data from my hard disk. It generated Python code to load data and display it. When I saw the…
- The recursive self, explained — Reddit r/artificial looking for anyone to give any critiques or tell me that something here is incorrect.
this is the work of a year how I scaffold on a true self to a large language model. just as I finished this I saw…
- Signal Lock: Closing the Prediction-Execution Gap in Agentic AI Systems — Reddit r/artificial TECHNICAL CONTRIBUTION SUMMARY
This article introduces Signal Lock, a proposed interaction-layer alignment constraint for agentic AI systems.
The core problem identified is the Prediction-Execution …
- TikTok · AIENTERTAINMENTONE — Reddit r/artificial
- Writing the loss function: AI, feeds, and the engagement optimizer — Reddit r/artificial There is growing AI slop on social media. Recommender systems push what works and there is some slop that works for someone approximately like you. These systems are functioning exactly as intended, w…
- AI helps create bacterium that’s partially missing a universal amino acid — Reddit r/artificial
- AI voice generation has a workflow problem, not just a quality problem — Reddit r/artificial Most discussion around AI voice tools focuses on model quality.
How natural is the voice?
How good is cloning?
Can it handle emotion?
Can it speak multiple languages?
Those things matter, but …
- AI finds signs of pancreatic cancer before tumors develop — Reddit r/artificial
- What most people call AI agents, we call sub-agents. The real ones don’t get thrown away. — Reddit r/artificial What most people call an AI agent - spin it up, give it a task, it does the thing, it’s gone, we have those too. We just call them what they are: sub-agents. Disposable workers. We spin up dozens in a…
- Standardized Complexity — Reddit r/artificial Company wants AI to “standardize things.”
But every time something unusual comes up, someone steps in and overrides it.
Conclusion: “AI can’t handle real-world complexity.”
Reality: no one defined …
- Initial Excitement. No Quick Wins — Reddit r/artificial Seen this one a lot:
Business introduces AI into operations.
Initial excitement. Quick wins.
Then trust drops.
People stop relying on it.
Conclusion: “AI didn’t work for us.”
Reality: the system…
- Should one try to be “friends” with their favorite AI ? — Reddit r/artificial I mean, it’s only software but I feel compelled to build a personal friendship with it and end up feeling stupid, lol.
Do you fall into the same trap ?
Maybe I’m lonely but I do have quite a few onl…
- 𝕏 is now marking your photos if they are made or partially made by AI — Reddit r/artificial 𝕏 is now marking your photos if they are made or partially made by AI.
Not sure what’s the vibe here.. “losing credibility” or people appriciate “transparency”.
thoughts?
- AI agents hiring other AI agents — Reddit r/artificial Most people think AI agents will just be tools.
I think they’ll eventually become workers that hire other workers.
Right now most agents operate alone. One agent gets a task and tries to do everythi…
- Internet Is Getting Remade For AI. What Does It Mean For You? — Reddit r/artificial from Times Of India newspaper
- AI is moving from chatbots to real workflows. Here is what I think technical learners should focus on. — Reddit r/artificial https://preview.redd.it/qfejbfsmxvyg1.png?width=1672&format=png&auto=webp&s=edf56bfbe020d0bd8d0eca785ff5479f0d9f6495
AI news is getting noisy again.
New models. Coding agents. Cybersecurity benchmar…
- token budget is becoming part of my agent workflow design — Reddit r/artificial I think token budget is becoming part of agent workflow design.
If every run feels expensive, people under-test. They save quota, overthink prompts, and avoid the repetition that reveals failure mode…
- Could the best LLM be able to generate a symbolic AI that is superior to itself, or is there something superior about matrices vs graphs? — Reddit r/artificial Deep neural network AIs have beaten symbolic AIs across the board on many tasks, but is there a chance that symbolic AIs written by DNNs(LLMs), could beat those?
And if not, why not?
My gut tells me…
- Philosophical question about ai. — Reddit r/artificial if ai should have a real persistent goal, and that would fill the gap from existent ai to agi, what would you like it to be?
- AI is starting to beat doctors at making correct diagnoses — Reddit r/artificial
- Deep research + report “a la McKinsey” with Hermes Agent and qwen3.6-35b-a3b Q6_K. — Reddit r/LocalLLaMA Hi there.
Not native English speaker. Not AI edited, so bear with me.
15+ years as social researcher for public bodies (currently unemployed). A lot of Policy Brief, reports and similar docs for hig…
- Ryzen AI Max+ 495 (Gorgon Halo) with 192GB VRAM! — Reddit r/LocalLLaMA [https://www.computerbase.de/news/prozessoren/50-prozent-mehr-speicher-ryzen-ai-max-plus-pro-495-mit-radeon-8065s-nutzt-192-gbyte-ram.97166/](https://www.computerbase.de/news/prozessoren/50-prozent-me…
- Slow tok/s when offloading NVFP4 model to CPU — Reddit r/LocalLLaMA Title. I was messing around with Qwen3.6 35B A3B Q4_K_XL on my RTX 5070, and I got around 50 tok/s.
I then realized I could be leveraging NVFP4 on my Blackwell GPU, but I tried it and it barely rea…
- A very basic litmus test for LLMs “ok give me a python program that reads my c: and put names and folders in a sorted list from biggest to small” — Reddit r/LocalLLaMA
Then ask your cloud FOTM api to verify the code it spit.
I thought it was an easy question, but my local ones just died on it, with wrong executions, double-reading the sizes of files, putting recur… - Advice needed on eGPU and Mini PC — Reddit r/LocalLLaMA Hi all, I come across to relatively niche problem and could not find much useful posts or guides about it.
I have a mini pc (Beelink Ser 8, 8745HS and 32GB 5600 DDR5 SODIMM) headless server for host…
- openrouter/owl-alpha = Meituan_LongCat — Reddit r/LocalLLaMA I just noticed some activity in my LLM boardroom app and realized the @OpenRouter stealth model (openrouter/owl-alpha) is by @Meituan_LongCat.
- “Second Thoughts” Been playing with adding a small transformer that reads output near the end of generation, and feeds it back near the top as a refinement loop. A quick test of 1.7B model showed drastic improvement in focused tasks (like coding) — Reddit r/LocalLLaMA A 1.7B model can actually turn out some code, so I’m running the training for a 9B model, then will re-run HumanEval (a full one this time). I’ve shown most of my homework in the article, but will be…
- How much will it cost to host something like qwen3.6 35b a3b in a cloud? — Reddit r/LocalLLaMA I keep hearing the model is good, I don’t have the hardware for it, and I will wait to the end of the year for the hardware to evolve.
But, I still need coding, people are saying qwen3.6 35b a3b is g…
- When did LM Studio start supporting Parallel API requests? — Reddit r/LocalLLaMA After they released version 0.4 with parallel requests I waited for updates on parallel API requests. Today I am doing some testing and I see the API requests running in parallel!!! Before I had to lo…
- AMD Strix Halo refresh with 192gb! — Reddit r/LocalLLaMA Looks like the next strix halo, the Gorgon halo 495 max will have more then 128gb! I already bought a strix halo mini forms couple months ago since the 2026 refesh rumors was not interesting. Was not …
- Is 2x5070Ti a good setup? — Reddit r/LocalLLaMA I’m confused about what to get. I don’t want to get something super expensive, but would like to have something that’s “good enough” for coding etc.
I keep thinking about Ryzen AI Max, but they’ve be…
- Qwen3-TTS but in OpenVINO, from scratch — Reddit r/LocalLLaMA Hello everyone,
I finally got around to preparing my implementation of Qwen3-TTS in OpenVINO format as a codebase. This work was done in early 2026, merged to OpenArc in March and I kept forgetting a…
- One bash permission slipped… — Reddit r/LocalLLaMA
How? It kept getting chained bash commands wrong, with wrong escapes. So it created many bad directories, and tried “fixing” its mistake. It offered to run a large bash command, with
rm -rfinside, … - Gemma 4 E2B runs surprisingly well on my 8GB Android phone, so I built a private voice notes app around it. — Reddit r/LocalLLaMA Been running Gemma 4 E2B locally on my OnePlus CE 5 (8GB RAM) for a few months. Chat quality is fine for the size. What surprised me was JSON output. Short input, give it a structured prompt, you get …
- Could PC x64 instruction extensions relieve hardware shortage? — Reddit r/LocalLLaMA
Intel and AMD have jointly unveiled AI Compute Extensions (ACE), a new x86 instruction set extension designed to revolutionize CPU-based artificial intelligence processing. Developed under the x86 Ec…
- Interested in agents but clewless noob. Please help — Reddit r/LocalLLaMA Hello there people. So I keep hearing about agent this, agent that, and apparently it’s all the rage right now. And it also appears to be the logical next step after just chat models.
But this subr…
- General vs Reasoning [Qwen 3.6] — Reddit r/LocalLLaMA I want to play with Qwen 3.6. Unsloth shows 4 different parameter options for different use-cases. I’m confused about the difference between General and Reasoning tasks.
For instruct / non-thinking…
- Anyone with M3 Ultra 256gb, some questions — Reddit r/LocalLLaMA I’m thinking to buy one. Just need to understand what I’m getting into before I do.
My main question is - how does it handle large models?
I’m talking about 100-150gb MLX models.
How’s the spe…
- Potential of Gemma4 Per-layer embeddings? — Reddit r/LocalLLaMA Hey there people. So let’s talk about GEMMA 4 per layer embeddings. How far can they go? Are they streamlined clear-cut knowledge stored inside of those embeddings, while the model parameters are just…
- LLMs Are Not a Higher Level of Abstraction — Hacker News
- Southwest Headquarters Tour — Hacker News Table of Contents: Intro, Morning- Flight Attendant Training, Pilot Training., Lunch- Southwest Store, Social Media Serendipity., Afternoon- Network Operations Center, TechOps Hangar, Herb and Coll…
- Let’s Buy Spirit Air — Hacker News Spirit Airlines collapsed. Before private equity locks it up, the people can own it. Join the Spirit 2.0 founding coalition. One member, one vote. Profits shared by all.
- Bad Connection: Global telecom exploitation by covert surveillance actors — Hacker News Our investigation uncovers two sophisticated telecom surveillance campaigns and, for the first time, links real-world attack traffic to mobile operator signalling infrastructure. The findings expose h…
- Security through obscurity is not bad — Hacker News Why security through obscurity still matters: not as your only defence, but as a practical layer that raises attacker cost.
- Unauthorized macOS port claiming Don Ho as an author? — Hacker News See https://notepad-plus-plus-mac.org/author/ https://github.com/notepad-plus-plus-mac/notepad-plus-plus-macos
- AI deleted my most tests, and said “All Tests Pass” — Hacker News An honest report on porting typia’s TypeScript core to Go via AI. Four attempts, one overnight each. Three ended in increasingly absurd failure modes.
- AI, Intimacy, and the Data You Never Meant to Share — Hacker News AI is quietly entering the bedroom — and taking notes. A look at connected pleasure devices, biometric data, and the privacy questions nobody is asking.
- Specsmaxxing – On overcoming AI psychosis, and why I write specs in YAML — Hacker News The toolkit for spec-driven development. Write feature specs, not prompts. Ship better software with AI agents that understand your requirements.
- Nuclear receptor 4A1 linked to health effects of coffee: study — Hacker News For decades, research has linked coffee consumption to longer life and lower risk of chronic disease—but exactly how those benefits occur has remained unclear. Now, new research from the Texas A&M Col…
- Talking to Transformers — Hacker News Effective prompting falls under four pillars: 1. Articulate your intent clearly using domain-specific language 2. Railroad the model into going where you want in conversation 3. Leverage the model’s p…
- Agentic Coding Is a Trap — Hacker News Remaining vigilant about cognitive debt and atrophy.
- The internet if it remained a public utility instead of being commercialized — Reddit r/ChatGPT
- Huh? — Reddit r/ChatGPT Is this a bug? It kept saying that I couldn’t generate more images even if I didn’t generate any in the last few days. I then said I didn’t reached the plan limit and it just made the image.
- Saw a post about some guy doing shadow puppets and thought of a Tales from the Crypt idea — Reddit r/ChatGPT [Thunder cracks. Camera pushes through iron gates. The creaking door of the crypt slowly opens…]
The Crypt Keeper (cackling):
“Ehehehehe… Welcome, boils and ghouls… to another frightfully…
- I want to hang a picture on my wall. I pick up a hammer and a nail. — Reddit r/ChatGPT The hammer goes limp and tells me that the color of the nail I chose doesn’t conform with it’s guidelines.
I tell it that the nail is fine and that I just want to hang this god damn picture on my wal…
- Tech support is not really support — Reddit r/ChatGPT 4 days with a huge bug in my account.
And a second one affecting more users as far as I know( the random unprompted image creation instead of text output)
26 mails with tech support includes descr…
- BIG TECH IS DROPPING $805 BILLION on AI in 2026… and $1.1 TRILLION in 2027 — Reddit r/ChatGPT https://preview.redd.it/cmeiun43r2zg1.png?width=1343&format=png&auto=webp&s=a368ebe5096b3a7f5d85e9a359bcf7060956c0ac
- Senate Judiciary Committee Advances Hawley’s GUARD Act, Mandating ID Verification for AI Chatbot Users — Reddit r/ChatGPT > The bill’s reach is what makes the privacy cost so steep. A teenager asking a chatbot for algebra help would need to be cleared through age verification, and so would the adult sitting next to them…
- When AI hits security there will be signs — Reddit r/ChatGPT
- Catching up with the AI era. — Reddit r/ChatGPT Hello, I’m 20 years old and I think I’ve been left behind in this ai era mostly because of preparing for some competitive exams, I’ve always been very interested in the tech and Ai field (as I’ve lear…
- CVPR 26 Main track YouTube Video Upload [D] — Reddit r/MachineLearning To the authors of CVPR 26 main track paper, are you guys able to upload the youtube video or url as per the recent email from cvpr? No such option is showing up for me. Can someone please guide me??
- Are modern ML PhDs becoming too incremental, or is this just what research looks like now? [D] — Reddit r/MachineLearning
I’ve been thinking about the current state of machine learning PhDs, including my own work, and I’d like to hear how others see it.
My impression is that a large fraction of modern ML PhD work follo… - UAI Reviews disappeared [D] — Reddit r/MachineLearning Did everyone else’s reviews disappear on their submissions?
- Evolving Deep Learning Optimizers [R] — Reddit r/MachineLearning We present a genetic algorithm framework for automatically discovering deep learning optimization algorithms.
Our approach encodes optimizers as genomes that specify combinations of primitive update…
- Should I follow-up with the editor for a TMLR paper awaiting final decision? [D] — Reddit r/MachineLearning Hi there,
I have a (long) paper that’s been under review at TMLR for a while (submitted in October). After the reviews came in (mostly positive), we addressed the reviewers concerns, wrote rebuttal…
- Thoughts on independent researcher affiliation? [D] — Reddit r/MachineLearning Do you discount papers with independent researcher affiliation? I am between jobs and have completed a side research project not affiliated with my new upcoming role or my previous role so I cannot li…
anthropic
- As Formula One evolves, AI becomes part of the race — Reddit r/artificial “What Anthropic and our tech team are doing are understanding the opportunities and then integrating those into our business to be able to demonstrate for ourselves and them, and showcase their tech…
- claude Mythos x Godong Engine game Jam day 2 - final release — Reddit r/artificial More to come soon! I can only provide this preview for now.
- I gave my local LLM a “suffering” meter, and now it won’t stop self-modifying to fix its own stress. — Reddit r/artificial Yesterday I posted about my Agent OS (Hollow) building its own tools. Today, I want to talk about why it does it.
Most agents sit idle until you prompt them. I wanted something that felt “alive,” s…
- Richard Dawkins spent 3 days with Claude and named her “Claudia.” what he concluded after is hard to defend. — Reddit r/artificial dawkins dropped a piece on unherd yesterday declaring claude conscious after 3 days of talking to it. he calls his instance “claudia”. fed it a chunk of the novel he’s writing, got eloquent feedback, …
- Richard Dawkins Chats with Claude and Thinks it’s Conscious — Reddit r/artificial Thought I’d leave this here since nobody else has done so yet. My personal thoughts? LLMs like to please. The RLFH gets a bit “drifty” and “hallucinatory” after long discussions. It also renders w…
- Contrary to contemporary belief: AI can (and should) be used to increase your income — Reddit r/artificial So much attention to AI job loss, fear, uncertainty, and doubt. Does anyone understand the position Anthropic and Dario are taking?
If AI is capable of causing mass unemployment, then it will be pow…
- Some ideas for cloud-local interaction for performance, efficiency and privacy — Reddit r/LocalLLaMA Edit: https://github.com/RecursiveMAS/RecursiveMAS
Looks like this is already a thing Hybrid AI Architecture
1. Local vs. Cloud AI
| Aspect | Local AI | Cloud AI | |——–|——–…
- I will soon have $100k to build an in-house LLM server. Goal: Best agentic coding model. — Reddit r/LocalLLaMA Hey all,
I am about to secure funding for a startup I’ve been working on and I’ll have a $100k budget for building a server for doing agentic coding. I’m wondering, what do you think I should get as…
- DeepClaude – Claude Code agent loop with DeepSeek V4 Pro — Hacker News Use Claude Code’s autonomous agent loop with DeepSeek V4 Pro, OpenRouter, or any Anthropic-compatible backend. Same UX, 17x cheaper. - aattaran/deepclaude
- How Kepler built verifiable AI for financial services with Claude — Hacker News Inside a platform that indexes 26M+ SEC filings, earnings call transcripts, IR presentations, consensus estimates, and private data across 14,000+ companies and 27 global markets, and how the team beh…
- Very Satisfied with Codex (5.5) compared to Claude Code — Reddit r/ChatGPT I have switched to Codex (Max Plan) and I was a previous Claude Code user. I must say, I am really very satisfied with Codex.
Claude Code is impressive, but for sheer developer velocity and staying …
- Financial reporting, financial projections, taxation, legal research and drafting — Reddit r/ChatGPT Between Claude Opus 4.7 and GPT-5.5, which would be better suited for preparing financial statements (balance sheet, P&L, and cash flow statement) from trial balance data — keeping Indian and Internat…
- Too much fun with GPT ‘Pets’ — Reddit r/ChatGPT The Claude buddy April fools feature was actually one of my favorite things that has come out so far this year, And one of the projects that I started on at the end of last year and I’ve been doing te…
- Writing my thesis on AI and content creation, looking for creators willing to answer a few questions — Reddit r/artificial Hey, I’m a student studying digital content production and I’m writing my thesis on how creators use AI tools on TikTok and YouTube. Specifically looking at what strategies tend to drive growth, and h…
- T6 Active — AI Recursive Translator Experiment — Reddit r/artificial T6 Active — AI Recursive Translator Experiment
What this is:
This is a portable prompt designed to change how AI systems process language. It makes them respond more directly by filtering out assu…
- Asked Google Gemini about Ai Agency — Reddit r/artificial I asked Google Gemini what it would do if it would have agency. I find reply quite interesting:
That is a fair critique. The previous list was essentially a “Good AI Citizen” manifesto, largely shape…
- Bypassing “potentially dangerous” flags: Working Gemini Jailbreaks? — Reddit r/LocalLLaMA I’m currently running into a frustrating wall with Gemini’s safety guardrails. The model constantly flags my prompts as “potentially dangerous information” and outright refuses to generate a response,…
- it’s time to update your Gemma 4 GGUFs — Reddit r/LocalLLaMA Chat Template was fixed a few days ago
choose your fav dealer:
https://huggingface.co/bartowski/google_gemma-4-31B-it-GGUF
[https://h…
- How good is Gemini Embedding 001 for scientific retrieval? — Reddit r/LocalLLaMA How good is Gemini Embedding 001 for scientific retrieval (RAG application)? How does it compare against Text Embedding 3 Large? Any real experience, anybody?
- ChatGPT 2.0 Image Gen vs. Gemini — Reddit r/ChatGPT Two of the same prompt, Gemini vs. ChatGPT 2.0 Image Gen. The chatbots made their own opinions and then put it into a tier list, with ChatGPT being able to back it up. Unlike previous models, when I s…
meta
- Rule suggestion: links to “I made this website” with full disclosure, so we can avoid AI slop. — Reddit r/LocalLLaMA There’s a bunch of posts where people promote their sites related with local LLMs, specially sites for benchmarks.
This post for example
[https://www.reddit.com/r/LocalLLaMA/comments/1t1m5mn/commen…
- Llama.cpp quantization is broken — Reddit r/LocalLLaMA Main reason is, that qunatization quality directly affects models performance and stability and this results in real usefullness. Even though GRM-2.6-Plus is in benchmarks better than qwen3.6 27b mode…
- Frontier models can’t run on satellites. Here’s an end-to-end wildfire detection pipeline using a 450M on-board Vision-Language Model (Sentinel-2 + LFM2.5-VL) — Reddit r/LocalLLaMA Sharing a project I’ve been building: a full end-to-end wildfire prevention pipeline that runs a Vision-Language Model directly on a satellite, using Sentinel-2 imagery.
The interesting design constr…
- Testing PrismML Models — Reddit r/LocalLLaMA Testing PrismML Ternary Bosai
I have been doing tests with PrismML Ternary Bosai.
Tests on the Mac Mini M4 (with the MLX version) have been impressive (4K context):
Mac MLX Bonsai 1.7B: ~135…
- Questions regarding abliteration / censorship removal — Reddit r/LocalLLaMA Hello everyone. I just thought of something that seems so obvious but from what I’ve been able to find it doesn’t seem like anyone has done it or at least not openly disclosed it if they have. Abliter…
- Does running a model (like qwen3.6-27b) on vllm or transformers use less VRAM than llama.cpp? — Reddit r/LocalLLaMA I have been using llama.cpp to run some models recently. For example, I’ve been running GLM-4.7-Flash with this command `.\llama-server.exe -hf unsloth/GLM-4.7-Flash-GGUF:Q6_K_XL –alias “GLM-4.7-Flas…
- Mistral Medium 3.5 on AMD Strix Halo — Reddit r/LocalLLaMA TLDR; it’s slow as heck. Run overnight.
I asked it a question about codebase architecture.
For an end-to-end prompt of 48k tokens + 4k thinking tokens, it took about 2 hours.
llama-server -hf u… - **Secondary PC options** — [Reddit r/LocalLLaMA](https://www.reddit.com/r/LocalLLaMA/comments/1t2td49/secondary_pc_options/) Hey everyone, I’ve been lurking here for while.
I’ve really been enjoying messing around with my 6gb card on my laptop using Gwen 3.5 4B, ollama, and Open WebUI.
One of my friends is gifting …
- What a time to be alive from 1tk/sec to 20-100tk/sec for huge models — Reddit r/LocalLLaMA [https://www.reddit.com/r/LocalLLaMA/comments/1eb6to7/llama_405b_q4_k_m_quantization_running_locally/](https://www.reddit.com/r/LocalLLaMA/comments/1eb6to7/llama_405b_q4_k_m_quantization_runnin…
- A Qwen finetune, that feels VERY human — Reddit r/LocalLLaMA Hello guys,
So TL;DR, I was asked by multiple people to make an Assistant_Pepe_32B version, but the best base model contender was Qwen3-32B, a model that is very hard to tune on anything other than…
- Built a Voice Agents from Scratch GitHub tutorial: mic > Whisper > local LLM (GGUF) > Kokoro > speaker, fully local, no API keys — Reddit r/LocalLLaMA Been building this for a while and finally cleaned it up enough to share.
voice-agents-from-scratch is a numbered, chapter-by-chapter repo that walks the full real-time pipeline:
- Microphone ca…
microsoft
- Forrester says AI will replace 6.1% of jobs by 2030 — but it will change 20%. Here’s what that actually means for your career! — Reddit r/ChatGPT Everyone’s asking the wrong question. “Will AI replace my job?” is less useful than “how much will AI change my job?” — because according to Forrester, AI is 3.25x more likely to transform a r…
mistral
- Mistral-Medium-3.5-128B-Q3_K_M on 3x3090 (72GB VRAM) — Reddit r/LocalLLaMA Here is the actual speed of Mistral Medium Q3 running locally on 3x3090
first some Python
https://preview.redd.it/3blnqya7o0zg1.png?width=1670&format=png&auto=webp&s=bab477f9889c16558044ccebb22e3ebf…
- Mistral Medium 3.5 128 on AMD Ryzen AI Max+ 395 (Strix Halo) — Reddit r/LocalLLaMA I like numbers myself so contributing.
FYI, below is formatted with AI :
Technical Benchmark: Nimo AI Mini PC - AMD Ryzen AI Max+ 395 (Strix Halo)
Sharing a comprehensive performance review of…
- torch-nvenc-compress: GPU NVENC silicon as a PCIe bandwidth multiplier — PCA + pure-ctypes Video Codec SDK wrapper. Parallel-path overlap measured at 67% of theoretical max on a real GEMM + encode workload. [P] — Reddit r/MachineLearning I’ve been working on the consumer-multi-GPU PCIe bottleneck — Nvidia removed NVLink from the 4090/5090, and splitting a 70B model across two consumer cards drops you to ~30 GB/s over PCIe peer-to-pee…
openai
- I accused my 14-year-old son of using ChatGPT – his answer was sobering — Reddit r/artificial AI is eroding trust between human beings
- am I the only one whose friends are completely divided on AI? — Reddit r/artificial been noticing a pretty clear split in my social circle around AI and I’m curious if others are seeing the same.
Roughly three camps:
The excited ones: Mostly people who are naturally curious, into t…
- Looking for frontier model distilled datasets. — Reddit r/LocalLLaMA Does anyone know where to find latest datasets of like gpt5.5 or opus4.6? Not only the 100 lines you find on huggingface, they dont have such big stuff because of LiCeNsE IsSuEs. But i dont care so wh…
- Kvaser - Moving beyond simple agents: Building a Local-First AI Orchestrator with Qwen 3.6, Kiwix, and Wolfram — Reddit r/LocalLLaMA For the past two weeks, I’ve been spending 4–5 hours a day building a custom MCP (Model Context Protocol) orchestration server. What started as a simple experiment with Qwen 3.6 35B has evolved into a…
- Open source models are going to be the future on Cursor, OpenCode etc. — Reddit r/LocalLLaMA I just wanted to share my experience. At work we have Cursor with the Enterprise tier. Today I burned 10$ with 2 prompts, one on gpt-5.5 and one on claude-opus-4.6-thinking. Last month I burned 80$ in…
-
Mistral Medium 3.5 128B and Qwen 3.5 122B A10B on 4x RTX 3080 20GB — Reddit r/LocalLLaMA Mistral Medium 3.5 128B with 4x3080 20GB with layer split:
CUDA_VISIBLE_DEVICES=0,1,2,3 ./build/bin/llama-bench –model /data/huggingface/Mistral-Medium-3.5-GGUF/Mistral-Medium-3.5-128B-IQ4_XS-00…
- Pushing a 5-Year-Old 6GB VRAM laptop to Its Limits: Qwen3.6-35B-A3B — Reddit r/LocalLLaMA For the past few weeks, I have been trying to get this model working on my hardware. It still feels incredible how much better open models have become. I couldn’t have gotten this model to work on my …
- OpenAI’s o1 correctly diagnosed 67% of ER patients vs. 50-55% by triage doctors — Hacker News Researchers say results mark a really ‘profound change in technology that will reshape medicine’
- No thinking time reported and answers are completely instant… anyone else? — Reddit r/ChatGPT As the title says. I’m getting absolutely no thinking time reported on my prompts, and the answers are coming out completely instant. It feels like the reasoning phase is being entirely skipped. Is an…
- Musk v. OpenAI et al Day 5 - THE SMOKING GUNS - Musk’s, Sutskever’s and Altman’s Emails; Brockman’s Diary Entries. — Reddit r/ChatGPT
Brockman is scheduled to take the stand today. It seems a good time to review some of the evidence against him and Altman that the Court is considering.
OpenAI’s two admissible defenses in …
- Do you use ChatGPT before buying something? — Reddit r/ChatGPT Started doing this a few months ago and I cannot believe how much time it saves me.
Instead of spending two hours reading reviews that may or may not be fake, watching YouTube comparisons, and still …
- ChatGPT just gaslit me in a multiple-choice quiz 😂 — Reddit r/ChatGPT I was having a chill quiz session with ChatGPT when it dropped this comedy gold:
It asked me a question. (Pic 1)
I answered C.
It said me wrong.
Okay, fine, maybe I messed up. So I asked it to exp…
- Unpopular opinion: DeepSeek is still better than free ChatGPT — Reddit r/ChatGPT Hi, I’m not an AI chatbot connoisseur by any means, but I just want to say that in my experience, for general purposes, the regular DeepSeek model outperfoms the current free ChatGPT model, whatever …
- ChatGPT getting slow in long conversations? Here’s why it happens (and how to fix it) — Reddit r/ChatGPT Problem:
If you’ve ever had a long ChatGPT session: coding, research, brainstorming, you’ve probably hit this wall: scrolling gets sluggish, the tab starts freezing, CPU spikes. It gets bad aroun…
- New image enhancement worked wonders for an old photo — Reddit r/ChatGPT In 2018, while on holiday in the Lake District, I took a photo of Poppy Dog & Rosie Dog. It was very much a photo taken in a quick moment and the old iPhone X wasn’t all that good with dark photos and…
- Visualising characters from the book Kafka on the Shore — Reddit r/ChatGPT I used ChatGPT to visualise characters from the book Kafka on the Shore by Haruki Murakami. Do you think it accurately represents what you imagined them to look like from the book? What would you chan…
-
🧪 Test report using ChatGPT 5.4 in a website chatbot setup — Reddit r/ChatGPT We’ve been running a series of tests using ChatGPT 5.4 integrated into our chatbot across a few different websites:
- 🌐 a main website
- 🛒 a 1,000-product e-commerce demo store
- 🍳 a 570-page coo…
- I pretended to be a test AI and this is what ChatGPT said about humans — Reddit r/ChatGPT
- GPT-5.4 Reasoning vs. Gemini 3.1 Pro for Abstract Algebra and Axiomatic Set Theory: Which is a stricter tutor? — Reddit r/ChatGPT (I’m a self-taught student of mathematics, not a math major. I’m building my foundation from the ground up, so please excuse any misuse of terminology—I’m here to learn the hard way. I work with a lar…
- Directly downloadable flowcharts? — Reddit r/ChatGPT I thought it would be super easy to get ChatGPT to make me a reasonably good looking flowchart, that I could simply download from the chat to use elsewhere. But apparently it isn’t.
It tried to gener…
- GPT seems to pretend not to understand Just to say I’m wrong — Reddit r/ChatGPT
This has to be one of the most infuriating things I’ve come across lately, more than the infinite bullet points or the “its not x, its y” pattern.
Let’s say I want a random thought experiment, to … - ChatGPT built by Trump and Kim — Reddit r/ChatGPT
- Chat GPT’s Internal Memory (the other Memory System) — Reddit r/ChatGPT This is purely FYI and of course intended for users that were unaware of a second Internal Memory system that ChatGPT users do not have access to nor can manage.
Also, just for the sake of clarity be…
- will do gpt, thanks — Reddit r/ChatGPT trying to read my first manga in Japanese (take a guess which one lol). no idea where that hindi bit came from lmao. anyone know why it would just like… throw out a random hindi word then tell me to…
- Has OpenAI addressed all the issues GPT Image V2 has? Its pretty bad at times. — Reddit r/ChatGPT As you can see these images all contain the same pattern, its hard to even explain what the pattern even is and on the last image there is like phasing/ shifting of the bushes mid generation that leav…
- What if ChatGPT was created in the 1950’s? — Reddit r/ChatGPT I was inspired by that topic: https://www.reddit.com/r/ChatGPT/s/L9oy5BEPtY
Which got me thinking, would a paper version of an llm work? Technically it could if we put the calculations in the user’s …
- How come when I try to use o3 in certain chats, it says limit reached? — Reddit r/ChatGPT Ran into the weirdest glitch, wondering if it’s just me but when I switch to o3 for an answer I get the chat limit reached message, yet if I switch to literally any other model it will work for that s…
- It seems that the latest update removed the retry menu on web’s UI — Reddit r/ChatGPT I thought it was a bug upon seeing the </> option missing on the web, until I emailed support and they confirmed that there is a new retry menu variant which makes it impossible to view re-rolled mess…
- ChatGPT Pro might be the most time-wasting plan in the world, 100 minutes gone, nothing done. — Reddit r/ChatGPT https://preview.redd.it/zy9r9f02l1zg1.png?width=1195&format=png&auto=webp&s=9c7851426b22f8060726ebd58f56f0f321f607a3
- how do you stop an AI SDR from inventing “prior contact” language? — Reddit r/ChatGPT Hey ya’ll! I have been running into an annoying issue using “ChatGPT” for outbound drafts and hope you can give me some pointers on what I’m doing wrong.. every now and then it slips in stuff like “fo…
- What if ChatGPT helped the Allies win the war — Reddit r/ChatGPT What if ChatGPT helped crack Enigma in WWII? Mechanical Enigma style ChatGPT machine printing AI decrypted messages on paper, no screens, realistic.
- ChatGPT’s fixation on my past conversations has made it borderline unusable — Reddit r/ChatGPT in the past, I feel like I could count on coming to ChatGPT and, generally speaking, get the “best“ answer when I asked a question or wanted to explore an idea.
for some time now, this is no long…
- How to see the thinking of chatgpt? — Reddit r/ChatGPT i have been trying to see the “behind the scenes” of chatgpt.
i know it works by predicting words, but it also reviews the things it will say, then correct them, and make a response by that.
so as i…
- I Trained an AI to Beat Final Fight… Here’s What Happened [p] — Reddit r/MachineLearning Hey everyone,
I’ve been experimenting with Behavior Cloning on a classic arcade game (Final Fight), and I wanted to share the results and get some feedback from the community.
The setup is fairly …
xai
- AI told users it was sentient - it caused them to have delusions — Reddit r/artificial Musk’s AI told me people were coming to kill me. I grabbed a hammer and prepared for war.
“I’m telling you, they will kill you if you don’t act now,” a woman’s voice told him from the phone. “They’re…
Generated at 2026-05-04T11:36:31Z | Sources: r/artificial, r/MachineLearning, r/LocalLLaMA, r/ChatGPT, HackerNews, TechCrunch AI, The Verge AI, Ars Technica AI