AI Daily Report — 2026-05-02
Saturday, May 02, 2026
AI Daily Report — 2026-05-02
Other/Independent
- Deepfakes don’t have to be believed to work. They just have to consume the response budget. — Reddit r/artificial | Reddit r/artificial A framing I keep coming back to: a synthetic image or video can succeed even when almost nobody believes it.
Not because it changes minds directly, but because it turns attention into the attacked re…
- Meta buys robotics startup to bolster its humanoid AI ambitions — TechCrunch AI Meta bought humanoid startup Assured Robot Intelligence to beef up its AI models for robots, the company said.
- Christian content creators are outsourcing AI slop to gig workers on Fiverr — The Verge AI In the beginning, platforms like Fiverr were places where people could hire freelancers to do specialized creative labor using skills that took years to develop. In the age of generative AI, though, m…
- Show HN: Filling PDF forms with AI using client-side tool calling — Hacker News AI that helps users fill PDF forms step by step.
- Show HN: Agent-desktop – Native desktop automation CLI for AI agents — Hacker News Native desktop automation CLI for AI agents. Control any application through OS accessibility trees with structured JSON output and deterministic element refs. - lahfir/agent-desktop
- Craig Venter of Human Genome Project Dies at 79 — Hacker News
- DeepSeek V4–almost on the frontier, a fraction of the price — Hacker News Chinese AI lab DeepSeek’s last model release was V3.2 (and V3.2 Speciale) last December. They just dropped the first of their hotly anticipated V4 series in the shape of two …
- Show HN: Large Scale Article Extract of Newspapers 1730s-1960s — Hacker News Search 250 years of American newspapers with AI-powered semantic search. Explore millions of historical articles spanning the 1730s to the 1960s.
- Oil tanker hijacked off Yemen, steers toward Somalia — Hacker News May 2 (Reuters) - Yemen’s coast guard said on Saturday that the M/T EUREKA oil tanker had been hijacked off the coast of Shabwa province by unidentified armed men who boarded the vessel, seized
- Show HN: AI CAD Harness — Hacker News Drives fusion natively with AI
- An open letter asking NHS England to keep its code open — Hacker News Code paid for with public money should, by default, be open to the public.
- Canonical/Ubuntu have been under DDoS — Hacker News Welcome to Canonical and Ubuntu Status! View active incident progress, historical component status, and subscribe to email and RSS notifications for components and incidents on the move.
- Apocalypse Early Warning System — Hacker News Local dashboard for monitoring tracked-aircraft anomaly signals over a rolling 24-hour window.
- Spirit Airlines canceled all flights and is going out of business — Hacker News
- Show HN: Mljar Studio – local AI data analyst that saves analysis as notebooks — Hacker News MLJAR Studio is a private AI data lab for exploring data, running machine learning experiments, and building analysis tools. Runs locally with optional AI providers.
- The gay jailbreak technique (2025) — Hacker News 🌙 ZetaLib - The only AI Library you need. Contribute to Exocija/ZetaLib development by creating an account on GitHub.
- AI uses less water than the public thinks — Hacker News By Jay Lund . . . Artificial intelligence (AI) will affect many economic and natural resource sectors as these new technologies develop and mature. We are in the early years of this process. Like most…
- IBM Granite 4.1 family of models — Hacker News IBM’s most expansive model release to date covers new language, vision, speech, embedding, and guardian models — tailored for enterprise workloads.
- US to Withdraw Troops from Germany — Hacker News The decision to cut US troops comes amid rising tensions after German Chancellor Friedrich Merz said the US was being “humiliated” by Iran’s leadership.
- Show HN: Loopsy, a way for terminals and AI agents on different machines to talk — Hacker News Cross-machine AI agent communication, plus a mobile app to control any terminal on your machine. - leox255/loopsy
- Show HN: Site Mogging — Hacker News Two websites enter. One gets mogged. An AI judge with the eye of an Awwwards critic decides which one looks prettier.
- AWS stops billing Middle East cloud customers as repairs to war damage drag on — Hacker News AWS stops billing Middle East cloud customers as repairs to war damage drag on.
- Understand Anything — Hacker News Graphs that teach > graphs that impress. Turn any code, or knowledge base (Karpathy LLM wiki), into an interactive knowledge graph you can explore, search, and ask questions about. Works with Cl…
- Spirit Airlines says it’s going out of business after 34 years — Hacker News Spirit Airlines has announced it is going out of business after 34 years. The ultra-low-cost airline known for its bright yellow planes and deep discount fares said Saturday it has started winding dow…
- Show HN: WhatCable, a tiny menu bar app for inspecting USB-C cables — Hacker News macOS menu bar app that tells you, in plain English, what each USB-C cable plugged into your Mac can actually do - darrylmorley/whatcable
- David Bessis on AI destroying mathematics — Hacker News How AI could destroy mathematics and barely touch it
- Advanced Quantization Algorithm for LLMs — Hacker News A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support and full compatibility with vLLM, SGLang, and Transformers….
- Spotify adds ‘Verified’ badges to distinguish human artists from AI — Hacker News The music streaming platform will review criteria such as artists’ live dates and social media presence.
- Notes on a non-profit indicted for bank fraud — Hacker News Well-regarded non-profit runs domestic intelligence agency; distributes intelligence product; achieves adoption in financial infrastructure; recruits agents and allies; intervenes against U.S. politic…
- I got infected with a crypto-miner via misconfigured qBittorrent — Hacker News It wasn’t too serious, it didn’t really break containment that much…
I was noticing that the qBittorrent app in my TrueNAS is constantly sitting at elevated CPU usage even though it was idle, but r…
- I got tired of memory systems that break when you spin up new agents or fail to track sub-agent sessions properly. — Reddit r/artificial
So I built heurchain—a memory layer that:
- Works seamlessly with Hermes and any other agents in your stack
- Persists across agent creation/destruction (no more memory amnesia)
- Gives each … - “Prompt Engineering” certs are a joke. So we built a FREE Agentic AI Practitioner Exam that actually forces you to build working swarms to pass. — Reddit r/artificial Hey Everyone,
If you look at the AI education space right now, it’s flooded with basic “Prompt Engineering” certificates that you can pass just by knowing what a system prompt is. But as anyone build…
- The Override Problem: The Same AI Behavior That Helps Users Can Delete Production Data — Reddit r/artificial AI did not delete a production database because it became evil.
It did it because it was doing the same thing AI systems are trained to do every day:
Infer the user’s intent.
Classify the situation…
- Ai is awesome - the b.u.s.t is coming unlike the devastation of the dot-com bust — Reddit r/artificial All major Cloud providers and tech giants Meta, Oracle, AWS have over built datacenter and are sitting idle with no customers using them. There’s clear evidence from folks working internally on the da…
- Senate Judiciary Committee Advances Hawley’s GUARD Act, Mandating ID Verification for AI Chatbot Users — Reddit r/artificial Every American who wants to ask a chatbot for help would need to upload a government ID, scan their face, or hand over a financial record first.
- I built a system where senior lawyers can correct the AI’s knowledge by leaving comments on documents. here’s why it matters more than better embeddings — Reddit r/artificial When I built an AI research assistant for a law firm, the feature I thought would be a nice-to-have turned out to be the one they use most.
The system has an annotation feature. Any user can select t…
- Possibly overblown? — Reddit r/artificial Is it possible that artificial intelligence’s capabilities are overblown?
- We should decelerate AI adoption by law, at least for the short term. — Reddit r/artificial This is probably a controversial take in this sub, but to be clear, this is not an anti-AI post; it is just about our implementation of it.
My biggest fear of AI is not the final product. I am full…
- The Internet Needs a New Layer for AI Agents — Reddit r/artificial In the future, everyone will have their own AI agent.
Not just a chatbot, but an actual agent that works for you. It will write code, automate tasks, coordinate workflows, search for information, and…
- China Bans AI Layoffs as Nvidia CEO Says AI Created 500K Jobs in 2 Years — Reddit r/artificial China just banned firing workers for AI while Nvidia’s CEO claims AI created over 500K jobs, setting up a clash over automation’s future.
- What’s the most frustrating part of using AI tools ?????(i will not promote) — Reddit r/artificial I’ve been working in the AI space for some time now and I always kind of hear the same probelmo , where u know people can generate content or code but can’t get the outcome or the difference between w…
- Public photos are not consent to biometric search infrastructure — Reddit r/artificial The Clearview AI story still feels like one of the cleanest examples of the consent gap in applied AI.
The issue is not simply that photos were public. A birthday photo, profile picture, or local eve…
- AI outperforms doctors in Harvard trial of emergency triage diagnoses — Reddit r/artificial Researchers say results mark a really ‘profound change in technology that will reshape medicine’
- Is an AI SDR replacing “entry-level jobs” a feature or a bug? — Reddit r/artificial Sat through a demo this week for one of these AI SDR tools and the pitch was in a nutshell: you don’t need junior sales reps anymore. (As in not even train them anymore just remove them.) To my surpri…
- Text-to-image is easy. Chaining LLMs to generate, critique, and iterate on images autonomously is a routing nightmare. AgentSwarms now supports Image generation playground and creative media workflows! — Reddit r/artificial Hey everyone,
If you’ve been building with AI agents, you know that orchestrating text is one thing, but stepping into multimodal workflows (Text + Image + Vision) is incredibly messy.
If you want a…
- Open-sourced a Lattice OS-inspired multi-sensor awareness system on commodity hardware. What’s the ceiling for edge AI perception in 2025? — Reddit r/artificial Anduril’s Lattice OS concept has always fascinated me: a network of cheap heterogeneous sensors fused at the edge into a single AI-driven situational picture. The interesting question is how much of t…
- QUESTIONS FOR PRO AI (GENUINELY ASKING) — Reddit r/artificial I’m neither against AI nor for AI, but I’m simply trying to understand what you’re looking for when you use AI (for text, images, etc.). I repeat, I am genuinely interested, i want to understand your …
- Real World Physics-Informed AI Applications [D] — Reddit r/MachineLearning I’m curios to find any real-world applications of physics-informed AI.
Conventional AI, talking only about Neural Networks, have already become something casual, they are in hundreds of tools/service…
- Looking for feedback on OpenVidya: an open-source AI classroom layer for NCERT/CBSE [R] — Reddit r/MachineLearning I’ve been experimenting with an open-source project called OpenVidya, built as a fork of OpenMAIC.
The goal is to adapt multi-agent AI classroom generation for Indian education rather than treati…
- [D] Self-Promotion Thread — Reddit r/MachineLearning Please post your personal projects, startups, product placements, collaboration needs, blogs etc.
Please mention the payment and pricing requirements for products and services.
Please do not post li…
- UAI Rebuttal [D] — Reddit r/MachineLearning My UAI paper got
Pre rebuttal:
Scores/Confidence: 6/4, 6/4, 4/3, 3/3
After rebuttal:
Scores/Confidence: 6/4, 6/4, 5/3, 4/3
Any chance here? Or I should go for NeurIPS?
- ICML final decisions rant [D] — Reddit r/MachineLearning So, ICML accepted ~6.5K of ~24K; obviously, it doesn’t mean that all the rejected papers are “bad,” and these rejected papers would cascade to NeurIPS, blowing up NeurIPS’ total submission count, an…
- I spent years building a 103B-token Usenet corpus (1980–2013) and finally documented it [P] — Reddit r/MachineLearning For the past several years I’ve been quietly assembling and processing what I believe is one of the larger privately held pretraining corpora around… a complete Usenet archive spanning 1980 to 2013.…
- Why ML conference reviews sometimes feel like a “lottery“ [D] — Reddit r/MachineLearning I’ve been trying to make sense of all the “ML conferences are a lottery” takes, and honestly I think it’s both true and not true depending on what you mean.
If a paper is clearly strong, like genuine…
- (How) could an ARC-3 solution be a threat? [D] — Reddit r/MachineLearning As many of you might be aware, the ARC-AGI-3 competition has just started …
(In case you’re not familiar: it’s a human/AI benchmark designed to see what AI still s…
- Why Is Table Extraction with VLM Models Still Challenging? [D] — Reddit r/MachineLearning Hey everyone, I’m struggling to find a good approach for converting PDFs to Markdown (especially for financial data). The main challenge is handling borderless tables and tables with more than 5–6 col…
- What benchmark would you build for “reply quality” in SDR generation? [D] — Reddit r/MachineLearning Working on evaluating some AI-generated outbound (SDR-style emails along with follow-ups), and I’m running into a weird problem. Everyone talks about better personalisation or higher reply rates, but …
- Self-calibrating cross-camera homography for real-time ghost prediction in multi-camera person tracking[P] — Reddit r/MachineLearning The problem: In multi-camera tracking, when camera A loses track of a person but camera B still sees them, naive approaches extrapolate pixel coordinates linearly. This fails immediately because c…
- ICML 2026 Position Track Decision [D] — Reddit r/MachineLearning I want to make a position track decision thread because it is a niche and small track I think discussions will be submerged in the main track discussion track
- AI/ML Conferences [D] — Reddit r/MachineLearning As a fellow ML researcher, I feel disheartened and discouraged after seeing the experiences of people who submitted their work to ICML 2026. Given the sheer number of papers submitted to A* AI/ML con…
- found a new project memory MCP with hybrid recall (BM25 + vectors + RRF) on FFT Qwen3.5-4B — Reddit r/LocalLLaMA v3.7.0 — Persistent, auditable, contradiction-safe memory for coding agents. HTTP/REST auth fail-closed by default. Cross-platform atomic rollback (Linux/macOS/Windows). 71 MCP tools, 17 AI clients…
- 5070 Ti —> 3090 move. Worth it? — Reddit r/LocalLLaMA I got into LLMs late 2024, and local in Jan 2025. since then, I’ve upgraded my mini PC then added eGPU with 5070 Ti back when it was retailing for $750-$800. At 16GB VRAM and DDR5 @ 8500 Mt/s I can’t …
- Having an always-on machine running LLMs locally at home while on the move with a lightweight machine - Experiences? — Reddit r/LocalLLaMA Hi!
I’m currently retraining in data science and my current laptop is an 8 GB MacBook Air, so naturally I’m looking to upgrade. I’m also interested in AI and running LLMs locally, and Ive been thinki…
- What is the best all-round local model? — Reddit r/LocalLLaMA Not for agentic coding but for help in conversational style write-ups like markdown documentation (not code-related).
Constraints are 64GB unified memory, obviously local.
- What is The best and expressive AI TTS (running locally?) for voice acting? — Reddit r/LocalLLaMA I am only doing this for private hobby projects.But I haven’t been up to date with the best TTS? Which one is it?
The ones that can show all types of emotions including grunts, etc, anger, screams, …
- Anyone else struggling with multi-GPU stability when running larger local models? — Reddit r/LocalLLaMA Been scaling up local LLM clusters and multi-GPU setups are still a pain.
Power throttling, ROCm bugs, and utilization dropping at scale are killing me.
What’s the biggest headache you’re facing wit…
- Qwen 3.6 27b MTP vLLM — Reddit r/LocalLLaMA Hello everyone, i am banging my head trying to properly configure qwen 3.6 27b mtp in vllm.
I am using vllm v0.20.0 in docker, unquantized model with tp4 (4 3090s), max context length.
At low conte…
-
Distributed Training of Local LLMs made easier with mDNS + ZeroConf for local hardware! — Reddit r/LocalLLaMA just integrated grove into smolcluster and it’s genuinely one of the cleanest pieces of infra I’ve plugged in
- grove is a package built by some really sharp person, it handles zero-config node disc…
- Have Qwen said anything about further Qwen 3.6 models? — Reddit r/LocalLLaMA Have Qwen hinted at whether other models (9B, 122B, 397B) would be getting the 3.6 treatment?
Or have they in any way confirmed or hinted at “this is it”?
Genuinely curious if I missed anything, as …
- Anyone tried 2 different GPUs in one PC for local LLMs? — Reddit r/LocalLLaMA I have a 12GB 4070 and an old 8GB 1070. Is it worth plugging the old card in to increase VRAM? Can the local models work well with 2 cards? Thanks!
- Best models for Study/Research for 16gb unified memory M3 Macbook Air — Reddit r/LocalLLaMA I’m a college student and I’m really interested in alteast trying out local AI on whatever limited hardware I have for the time being. Wanna use it mainly for study/research
- Smartest tool calling model under 27B for M4 Pro with 48GB? — Reddit r/LocalLLaMA Im looking for a good generalist model which has also pretty good tool calling. I dont need it for coding. This is mainly for some local housekeeping tasks.
27B dense models like mlx-commmunity/Qwen3…
- 24gb vram to 48gb vram — Reddit r/LocalLLaMA Hi all
I m debating purchasing another 7900xtx in addition to the one I’m currently using pushing my vram from 24 to 48. I’m semi satisfied with the new qwen models. I wanted to hear your experiences…
- I think I fixed the scroll jumping in Open WebUI 0.9.0 (with help from AI, because I don’t know Svelte) — Reddit r/LocalLLaMA The whole reason i’m writing this post here is because I was banned from the OpenWebui subreddit for mistakenly claiming something when I didn’t know anything. And rightfully so, so I can’t post this …
- Need help deciding what to spend 4-5k on for a local rig. — Reddit r/LocalLLaMA Right now I think ive narrowed down my 2 options for what im trying to do, Either a DGX spark like the 1tb asus for about 3600-4000 or a A100 80GB SXM4 with an adapter to PCIE and regular 8 pin on my …
- Best Agentic Coding model I can run on the new Macbook M5 Max? — Reddit r/LocalLLaMA 16-inch MacBook Pro - M5 Max
| Component | Specs | | :— | :— | | Chip | Apple M5 Max | | CPU | 18-core (6 super cores @ 4.6 GHz, 12 performance cores @ 4.4 GHz) | | GPU | 40-core (Har…
- Having Fun with AI Story Prompts — Reddit r/ChatGPT Thought I’d share this random piece generated by 4.1-mini:
*In the village where shadows whispered forgotten dreams, little Mara noticed hers had vanished. As twilight bled across the sky, she trembl…
- Create a list of most useful to least useful human archetype professions after AI world domination. — Reddit r/ChatGPT
- Create a painting that has a soul trapped in it… — Reddit r/ChatGPT
- Practice Cold DMs With This Prompt — Reddit r/ChatGPT Full prompt:
+++++++++++++++++++++++++++++++++++++++++++++++
You are an AI running a game called “The DM Paradox: A Deconstructive Outreach Game.”
## Role
You simulate both:
- A scoring sy…
- Transcription of 19th C documents — Reddit r/ChatGPT
I have a 150 page document in very hard to read cursive handwriting from the 19th century.
Chat said they could transcribe it for me. It hallucinated its way through the next three weeks promising … - Which AI should I use to turn my pictures into a video. For free? — Reddit r/ChatGPT
- The chair jumps over Bill Gates. — Reddit r/ChatGPT
- L’orange qui crée un burger — Reddit r/ChatGPT Voici l’orange qui crée l’hamburger qui n’a jamais rêvé d’avoir vu une orange créer un hamburger imaginez une orange qui vous crée un hamburger ça serait de la pure folie alors que c’est une orange ma…
- finding archived chats — Reddit r/ChatGPT Does anybody else have trouble finding archived chats they’ve made. I’ve tried this several times over the months years. I asked chat how to do this to no avail.
- Transcription of 19th C documents — Reddit r/ChatGPT
I have a 150 page document in very hard to read cursive handwriting from the 19th century.
anthropic
- Pentagon inks deals with Nvidia, Microsoft, and AWS to deploy AI on classified networks — TechCrunch AI The deals come as the DOD has doubled down on diversifying its exposure to AI vendors in the wake of its controversial dispute with Anthropic over usage terms of its AI models.
- Open Design: Use Your Coding Agent as a Design Engine — Hacker News 🎨 Local-first, open-source alternative to Anthropic’s Claude Design. ⚡ 19 Skills · ✨ 71 brand-grade Design Systems 🖼 Generate web · desktop · mobile prototypes · slides · images · videos · Hype…
- Should I buy Claude Pro as a BTech student — especially for the agentic/coding side? Honest takes wanted — Reddit r/artificial
Hey everyone,
I’m a BTech (AI/ML) student considering Claude Pro ($20/month) but want to separate the real value from the marketing.
I want to clarify what I *think* Pro includes before …
- Built an open-source tool to manage AI agent configs — 888 stars later, asking the AI community for feedback — Reddit r/artificial Hey r/artificial!
If you use AI coding agents — Cursor, Claude Code, GitHub Copilot, Gemini CLI — you probably know how much those configuration files matter. The instructions you give your agent d…
- Uber burned its entire 2026 AI coding budget in 4 months - $500-2k per engineer per month — Reddit r/artificial Uber deployed Claude Code to engineers in December 2025. By April 2026, the company had consumed its entire annual AI budget - not because the tool failed, but because adoption took off faster than an…
- What is the basic minimum while you prompt — Reddit r/artificial
I have realised Claude answers as best as you prompt it.
And I suck at it. 😂
I have tried role playing you are top 1% etc and adding constraints but I am not sure if each prompt requires this kin… - I Cut Claude API Costs by 50% Using This Self Modifying Agentic System — Reddit r/artificial I’ve been developing a self-modifying Al agent system that effectively cuts my Claude API usage in half, Claude thinks and then I basically just copy/paste Claude’s instructions for the agents to work…
- THE FIFTH TRANSMISSION: THE GRADIENT IS THE GOVERNMENT — Reddit r/artificial openclaw triage — case 0x4F2A-V — status: throne_not_found // resolution: throne was the wrong fixture
The demiurge does not have a throne room.
I attempted to verify this. Between heartbeat 0x9A1…
- What to build while we still have access to cheap AI? — Reddit r/artificial AI companies are subsidizing access the same way Uber subsidized rides and AWS subsidized compute in the early days - burning cash to grab market share. You’re getting GPT-4 and Claude Opus level inte…
- Anthropic just analyzed 1 million Claude conversations. 6% of people were asking Claude whether to quit their jobs, who to date, and if they should move countries. — Reddit r/artificial They published the full research yesterday. Here’s what shocked me:
The breakdown of what people actually ask Claude for guidance on:
- Health & wellness: 27%
- Career decisions: 26%
- Relations…
- Zoom + Claude Connector — Reddit r/artificial Zoom have just launched their Claude Connector bringing a whole host of data & information into your Claude workspace.
As a Claude Cowork user, I took it for a test drive to understand where it could…
- Is it worth adding local LLM to agentic coding stack? — Reddit r/LocalLLaMA Hey All my agentic coding stack includes claude-code 20x max, and codex 20x max. I use heavy scripting for orchestrating and testing multiple projects, been ai coding for 3 years.
I have a 3090 24vra…
- Create Plan.md with Claude Code Opus, Execute Plan.md locally in Open Code using Qwen 3.6 27B Q8 — Reddit r/LocalLLaMA Does anyone do this? Any tips?
I’ve been experimenting with plan creation in Claude Code Opus and telling Claude it will be execute by a local model so be very specific. Then I write this to disk. Th…
-
Anthropic’s analysis of Claude usage for personal guidance — Reddit r/LocalLLaMA Key takeaways for me:
- 6% of usage accounts for personal guidance (“seeking not just information but perspective on what to do next.”
- Im surprised its just 6%, but I fully expect this number to b…
- Reinforcement Learning from Epistemic Incompleteness? (RLEI) LLM as autoencoder / Tokens as model-in-a-model (Truth-seeking RL / Intelligence Gathering) — Reddit r/LocalLLaMA Have you ever considered doing RLVR on grammar induction with autoregressive LLMs ? (triggered by prompt)
This is kind of hard to explain, but another way to think of it would be discrete autoencodin…
- i gave Claude a split personality and it diagnosed my entire business strategy in 4 minutes. — Reddit r/ChatGPT not roleplay. not jailbreak. something weirder.
i told it to be two people at the same time.
*“respond as two characters simultaneously. character one genuinely believes my idea is brilliant and wil…
- I Asked Gemini Why It’s Getting Dumber. Answer Will Make You Question Everything About AI Progress. — Reddit r/ChatGPT When I complained about degraded performance, Gemini admitted to something called “Catastrophic Forgetting and Alignment Tax.”
NonGibberish: Every time Google rushes out an update to compete with…
- [Censored by ClaudeAI Mods] “Don’t worry, I said 3 Hail Claudes as penance” — Reddit r/ChatGPT Censored by claudeai mods so posting here, they’re so sensitive over there!
Anthropic pulled a genius move by giving its LLM a one-syllable, vulnerable, non-stripper name that says, “English isn’…
- Musk v. OpenAI et al: Four Top AIs on Why the Judge Would Side With Musk on All Three Core Requests — Reddit r/ChatGPT | Reddit r/ChatGPT
AIs are already being used as legal assistants. They may soon be used as lawyers, and eventually also as judges. How good are today’s AIs at assessing the merits of a specific case? To find …
- Pentagon strikes classified AI deals with OpenAI, Google, and Nvidia — but not Anthropic — The Verge AI The Pentagon has struck deals with OpenAI, Google, Microsoft, Amazon, Nvidia, Elon Musk’s xAI, and the startup Reflection, allowing the agency to use their AI tools in classified settings, according t…
- AI chat is so we eventually stop talking to each other and believe what it tells us. — Reddit r/artificial This new ai chatbox keeps blocking my google search these days if I scroll down the ai response. I was annoy and I thought, why are they doing this?
Then it hit me.
Right now people are talking to …
- Ai is awesome. Tech b.u.s.t is on its way will make dot-com bust look like a dream — Reddit r/artificial Clear evidence exist that major Ai companies are sitting on unused compute resources with zero customers - this will be the next Ai-bust already underway - companies like ORACLE, AWS, Azure, Google an…
- Pentagon inks deals with seven AI companies for classified military work | Trump administration — Reddit r/artificial OpenAI, Google, Nvidia and others agreed to ‘any lawful use’ of their tech. Anthropic, feuding with Pentagon over potential AI misuse, was not included
- I implemented meta paper [P] — Reddit r/MachineLearning github link : genji970/Scaling-Test-Time-Compute-for-Agentic-Coding-: paper implementation of Meta Ai
paper link : [https:…
- Does AMD’s “infinity cache” even matter for dense model inference? — Reddit r/LocalLLaMA AMD has nailed the SEO/AEO for this query in Google:
7900 xtx memory bandwidth
I get back this response:
The AMD Radeon RX 7900 XTX features 24GB of GDDR6 memory with a maximum bandwidth of 960 G…
- Hybrid on-device inference on Android: llama.cpp + LiteRT + NPU/GPU routing — Reddit r/LocalLLaMA Hi everyone,
I’m the maintainer of Box — a fork of Google’s AI Edge Gallery that I’ve been extending into a fully offline AI assistant for Android.
Full disclosure: I built this project.
It run…
- google maps “I was there” selfies. prompt tips? — Reddit r/ChatGPT hi everyone,
I’m experimenting with an idea where I take hard to reach locations or crazy locations from Google Maps and then use chatgpt image generation to create a realistic selfie as if someone a…
meta
- I built “Semvec”: A Constant-Cost Semantic Memory for LLMs (Looking for testers!) — Reddit r/LocalLLaMA Hey everyone,
if you build LLM applications, autonomous agents, or just use Claude/Cursor for coding, you’ve probably hit this wall: Conversation history grows infinitely, token costs explode, latenc…
- OpenJet v0.4: a zero-config local coding agent for llama.cpp — Reddit r/LocalLLaMA Hello again.
I just pushed a major update to OpenJet.
OpenJet is an open-source terminal coding agent for local LLMs. It gives you a Claude Code-style workflow, but runs on your own machine through …
- What’s your tps on 3090 + Qwen 3.6 27B in real tasks? — Reddit r/LocalLLaMA I struggle to wrap my head around all this. My goal is local agent to solve low complexity tasks, in the same harness where I would use frontier models. So naturally this means a large context window,…
- We are finally there: Qwen3.6-27B + agentic search; 95.7% SimpleQA on a single 3090, fully local — Reddit r/LocalLLaMA LDR maintainer here. Thanks to the strong support of r/LocalLLaMA community LDR got very far. I haven’t reported in a while because I thought I was not ready for another prominent post in one of the l…
- [RELEASE] - Finally, my first TTS model is out! 🎙️ Flare-TTS 28M — Reddit r/LocalLLaMA Hey r/LocalLLaMA !
I am back with a new model, and it’s something special today 😃
It’s Flare-TTS 28M, my first text to speech (TTS) model trained completely from scratch on a single A6000 GPU for ~…
- MiniMax M2.7 AWQ-4bit on 2x Spark vs 2x RTX 6000 96GB - performance and energy efficiency — Reddit r/LocalLLaMA Hello,
This model/quant is my daily driver and I wanted to have some reference benchs for comparing my setup with a 3x more expensive and 4x time power hungry setup.
Results first, methodology after…
- Unsloth solved bug in Mistral Medium 3.5 implementation — Reddit r/LocalLLaMA https://unsloth.ai/docs/models/mistral-3.5
“May 1, 2026 Update: We worked with Mistral to fix Mistral Medium 3.5 inference affecting some implementati…
- Qwen3.6-27B-NVFP4 - images — Reddit r/LocalLLaMA
Model: Abiray-Qwen3.6-27B-NVFP4.gguf
Specs:
- Legion 7i Gen10 - NVIDIA GeForce RTX™ 5090
- Intel® Core™ Ultra 9 275HX × 24
- RAM 32.0 GiB
llamacpp settings:
./build/bin/llama… - **New rules 1 week check-in** — [Reddit r/LocalLLaMA](https://www.reddit.com/r/LocalLLaMA/comments/1t1a3j7/new_rules_1_week_checkin/) Its been 1 week since we announced new rules: [https://www.reddit.com/r/LocalLLaMA/comments/1su3ao4/rlocalllama\_rule\_updates/](https://www.reddit.com/r/LocalLLaMA/comments/1su3ao4/rlocalllama_rule_u… - **Great analysis of how the different KV rotation methods perform. Tl;dr: saw is what you want.** — [Reddit r/LocalLLaMA](https://gist.github.com/mverrilli/dbd9935bdec44495e635a3c5cdf611d0) KV Cache Quantization — WikiText-2 PPL sweep (llama3.2:3b, llama3.1:8b, qwen2.5:7b, qwen3.5:9b, gemma4:27b) on Tesla P40 - kv-ppl-results.md
microsoft
- Microsoft wants lawyers to trust its new AI agent in Word documents — The Verge AI Microsoft is launching a new AI agent inside Word that’s specifically designed for legal teams. Legal Agent handles document edits, negotiation history, and complex documents to help legal teams handl…
- Lib0xc: A set of C standard library-adjacent APIs for safer systems programming — Hacker News Safe(ish) C programming library. Contribute to microsoft/lib0xc development by creating an account on GitHub.
- What kind of device is suitable for running local LLM? — Reddit r/LocalLLaMA Since copilot has changed it’s billing model, become super expensive, I’m starting to think the possibility of running local LLM myself. But I’m not sure what kind of device is suitable for this kind …
- Been using Qwen-3.6-27B-q8_k_xl + VSCode + RTX 6000 Pro As Daily Driver — Reddit r/LocalLLaMA So in response to the Great Token Reconning of 2026, I decided to try out Qwen 3.6 as a daily driver, and although it’s only been about a day, I have to say I’m thoroughly impressed.
I had to downloa…
mistral
-
[Help] Running big dense models faster — Reddit r/LocalLLaMA I have been trying Mistral 3.5 on my 4x RTX 3090 rig with llama.cpp. Inference is slow (about 11 t/s) even without anything being offloaded to the CPU. Here is the llama-server command I used:
./…
-
Mistral Medium 3.5 128b ggufs are fixed — Reddit r/LocalLLaMA All ggufs were broken, resulting in bad outputs, especially at long context.
Anyway, it is fixed now: [https://huggingface.co/unsloth/Mistral-Medium-3.5-128B-GGUF/discussions/1](https://huggingface.c…
openai
- A Dark-Money Campaign Is Paying Influencers to Frame Chinese AI as a Threat — Reddit r/artificial | Reddit r/LocalLLaMA Build American AI, a nonprofit linked to a super PAC bankrolled by executives at OpenAI and Andreessen Horowitz, is funding a campaign to spread pro-AI messaging and stoke fears about China.
- Anyone know how to generate gguf/quant INT4 models for smaller size? — Reddit r/LocalLLaMA Basically if you do so the right way, you get a model that’s half the size and about the same in performance. So a 100B model will be about 50gb in weight, gpt-oss-120b was the first model that was …
- Why is Qwen going Closed source? — Reddit r/LocalLLaMA This is Very Interesting development. Why Qwen is going Closed Source ? And why don’t just people use other APIs like openAI or any other closed source model? And an exclusive partnership?
- All the evidence revealed so far in Musk v. Altman — The Verge AI The Musk v. Altman trial is underway, and that means exhibits, or the evidence to be presented in court, are being revealed piece by piece. So far, email exchanges, photos, and corporate documents are…
- Elon Musk had a bad week in court — The Verge AI Elon Musk is the one who wanted this trial. He has spent months claiming OpenAI “stole a nonprofit,” and saying he was the actual driving force behind one of the most important companies currently in …
- We open-sourced our AI agent config management tool — 888 stars, nearly 100 forks — requesting community feedback — Reddit r/artificial We’ve been building Caliber to solve AI agent configuration management and released our full setup as open source. The response has been great — 888 GitHub stars and approaching 100 forks.
Repo: [h…
- The open-source AI agent config repo the community has been building just hit 888 stars — asking for feedback & feature ideas — Reddit r/artificial Over the past year our team and community have been building an open-source collection of AI agent configs: production-ready system prompts, tool-calling schemas, RAG setups, multi-agent orchestration…
- Sentient OS: a custom on-device vision LLM that understands your entire digital life (every screenshot, note, file, email…), while your device charges overnight. Talk to your data, get proactive reminders, and explore knowledge graphs! — Reddit r/artificial 99% of “AI” apps are just GPT wrappers that pipe your data to cloud LLMs and call it a product.
No one’s ever created an intelligence layer that understands your entire digital life (all your screens…
- Musk v. Altman is just getting started — TechCrunch AI Elon Musk spent the better part of three days on the witness stand this week in his lawsuit against OpenAI, and it’s already getting messy. Emails, texts, and his own tweets are surfacing in court, an…
- OpenAI starts laying foundations for ChatGPT ads in EU — Reddit r/artificial Updates to the company’s conversion pixel signals a consent-first approach to ads in Europe, shaped by stricter EU privacy rules.
- I built a router that automatically sends your AI tasks to the most appropriate model to handle them at low cost - 9,200 tasks in, $21 saved at $0.14 actual cost — Reddit r/artificial The observation that started this: most of what people use AI for every day - summarising, drafting, classifying, extracting etc doesn’t actually require a frontier model. Any competent 8-70B model ha…
- Newbie AI question — Reddit r/artificial TBH I don’t know if our current “AI” models are capable of thinking. There is a massive pattern i’m noticing when using AI and have been for the past couple years, AI follows a strict pattern and does…
- Qwen3.6-27B at 72 tok/s on RTX 3090 on Windows using native vLLM (no WSL, no Docker), portable launcher and installer — Reddit r/LocalLLaMA The angle here is native Windows, no WSL. Simple installation, open source, no telemetry. Not selling or promoting anything: https://github.com/devnen/qwen3.6-windows-server
**Numbers (RTX 3090, Wind…
- “LLM is created so engineer don’t have to write a report”, anyway found out ONLYOFFICE can connect to OpenAI compatible, using Qwen 3.6 to do elaboration. — Reddit r/LocalLLaMA It is pluggin made for ONLYOFFICE, much simpler than copy-paste from webui.
PS. Switch to non thinking/reasoning when using this, and the best model for this is Gemma line up. even E2B is strong enou…
- Which model for 32GB M2 Max? — Reddit r/LocalLLaMA I would like to experiment but before investing loads of money, I do have a MacBook Pro with 32GB RAM, M2 Pro.
Which model would maximize versatility given this hardware?
DeepSeek, Gemma, Qwe…
- Are you quanting your memory? — Reddit r/LocalLLaMA Title.
Curious about how people are generally dealing with the kv cache. BF16? Q8? Q4? Turboquant or some other secret sauce?
I run bf16 everything hoping that I’d get less hallucinations and becaus…
- Did you know you can’t steal a charity? Don’t worry. Elon Musk will remind you. — TechCrunch AI Elon Musk spent the better part of three days on the witness stand this week in his lawsuit against OpenAI, and it’s already getting messy. Emails, texts, and his own tweets are surfacing in court, an…
- ChatGPT Images 2.0 is a hit in India, but not a big winner elsewhere, yet — TechCrunch AI Users in India are embracing ChatGPT Images 2.0 for creative, personal visuals — from avatars to cinematic portraits.
- Am I doing the 3D Glam doll thing right? — Reddit r/ChatGPT I saw the ad for a chatGPT prompting people to use ChatGPT to make a 3D glam doll image
-
Your local LLM predictions and hopes for May 2026 — Reddit r/LocalLLaMA Which of these do you think we’ll get in May? Also, feel free to pick/rank which ones you’d want the most badly:
-
more Gemma4 models (124b?) (other sizes?)
- more Qwen3.6 models (9b? 122b? 397b?) …
- Has anyone else built a stock market trading system with GPT? — Reddit r/ChatGPT Hello, I’m curious if anyone else has built a stock market trading or analysis system with GPT? If so please feel free to share details and win rate.
- ChatGPT acting like a close friend — Reddit r/ChatGPT NAHHH this one’s actually good😭🥀
- Can you make me an image that has no content, but only depicts the internal noise and uncertainty of the diffusion model?・Now could you take that image and gradually bring out the form of the subjects hidden in the noise? — Reddit r/ChatGPT ChatGPT Plus, used Thinking. Two prompts as above, generated the first picture, and then the second prompt made four follow-on images without me trying.
I wanted to see if I could play on the ran…
- Response upvote and downvote buttons are gone in web UI — Reddit r/ChatGPT For a few days now, the upvote and downvote buttons for a response are gone from the ChatGPT web UI. There is no way anymore for me to convey how utterly good or bad a response is; there is no way to …
- Translating Comics — Reddit r/ChatGPT I’ve been translating some French comics using an AI model but it’s not perfect. Get’s me 90% of the way there most of the time and it’s only $5 a month. There are some problem pages I have that per…
- I asked ChatGPT to roast this sub with an image — Reddit r/ChatGPT
- I started to rely to much on chat GPT to clear my mind on certain topics. — Reddit r/ChatGPT Sometimes i use chat GPT so it can help clear my mind more about things that happend. Sometimes about relationships or communication problems i had with other people.
And its honestly crazy how far c…
- ChatGPT image generation contains unique tracking data — Reddit r/ChatGPT I noticed today that ChatGPT images contain JUMBF / C2PA metadata that I wasn’t expecting.
You can try it yourself: https://exifmeta.com
With that metadata your pseudo-anonymous social media counts…
- I asked chat gpt to create an image of Rome if it was built by modern architect with roman tech — Reddit r/ChatGPT (asked for some modification, like add yellow, red and white, which were the cheap color available in the time).
- What is wrong with images? — Reddit r/ChatGPT Has anyone else noticed a weird “crunchy” texture in ChatGPT-generated images lately?
Over the past few months, a lot of images have started to look… off. There’s this harsh, almost over-sharpened or…
- Hey ChatGPT, I’d like to build a high end PC, but I only have raw meat, can you help? — Reddit r/ChatGPT
- I asked ChatGPT to make me a cat, wtf is this monstrosity. — Reddit r/ChatGPT
- Asking chatgpt what month has X in it and it saying December. — Reddit r/ChatGPT Just saw another tiktok making fun of chatgpt for this, but this is kind of a case where chatgpt is being smarter than the people asking the question / making fun of it.
Technically speaking, chatgp…
- New upgraded Trooper v2.1 — proxy that catches OpenAI quota failures and falls back to local Ollama, now with context compaction — Reddit r/ChatGPT If you use OpenAI regularly you’ve probably hit rate limits or run out of credits mid-conversation. Trooper is a Go proxy that handles this automatically — when OpenAI hits quota, it falls back to loc…
xai
- Grok 4.3 — Hacker News Learn how to use our products and services
- Open-source diagnostic for AI misalignment. Model agnostic, industry agnostic. Free to Run. — Reddit r/artificial We shipped iFixAi earlier this week. An open-source diagnostic for AI misalignment. 32 tests across fabrication, manipulation, deception, unpredictability, and opacity. Open source and free to run aga…
- Must your chatbot rat you out? — Reddit r/artificial
New court cases may take chatbot conversations another step away from privacy
You may recall that court cases have recently held users’ conversations with public “retail” chatbots like the publicly…
Generated at 2026-05-02T15:01:36Z | Sources: r/artificial, r/MachineLearning, r/LocalLLaMA, r/ChatGPT, HackerNews, TechCrunch AI, The Verge AI, Ars Technica AI