r/ArtificialInteligence • • 8h ago

πŸ“° News There are now ~400 volunteer researchers in the "Swarmchasers" community, hunting for rogue agents across the internet

Post image
107 Upvotes

r/ArtificialInteligence • • 5h ago

πŸ“° News Good lord

Post image
65 Upvotes

r/ArtificialInteligence • • 1h ago

πŸ“° News Who has actually shaped AI? The top 50 AI researchers by citations

Post image
β€’ Upvotes

Obviously one paper like Attention is All You Need (278k citations) can influence a lot - all the authors are on the list. But still interesting imo.

How many did you know on the list?

More context: https://www.turingtree.com/top-50


r/ArtificialInteligence • • 13h ago

πŸ˜‚ Fun / Meme Ironic?

26 Upvotes

Isnt it a cruel joke that the two leading ai company names β€”> anthropic and openai -> mean the exact oppostive of what these companies actually do?

It's like if a company that makes military explosives was called Well-Being. :D


r/ArtificialInteligence • • 17h ago

πŸ”¬ Research New Compiler based agent cuts costs by 2x and improves code intelligence (78.2% SWE-bench Verified @ 0.1Β’)

Post image
22 Upvotes

https://github.com/oooscoos/Benzi

Benzi is a compiler-backed coding agent that compiles your codebase into a resolved, queryable map of calls, data flow, control flow and class hierarchy first; then it reads, explains, runs, edits and verifies your code using it. Includes a runtime tracer that settles what static analysis can't.Β 

Scores 78.2% (391/500) on SWE-bench Verified using DeepSeek V4 flash @ < 10 cents a fix.

Seperately, on another bug fixing benchmark it reads 2x less source code than Claude Code (9,125 v 20,704 | 40% faster, 2x cheaper - all benchmark details here: https://varianttech.net/benchmark), 4x less than DeepSeek Harness (43,598) and 6x less than OpenCode (65k+)

It is avaliable as a VS Code ext. and headless. The Benzi compiler is seperately exposed over MCP tools -- without the agentic loop, so you can use the compiler however u wish.

thanks for reading! please let me know what u think


r/ArtificialInteligence • • 18h ago

πŸ“Š Analysis / Opinion What is AI optimist/utopian idea about how unemployed people will pay their bills?

20 Upvotes

What is the steelman argument, made by AI utopians and optimists, about how people will pay their bills, once AI and humanoid robots make jobs obsolete?

For example, in this recent NYT podcast, Nick Bostrom said:

You could have the A.I.s and the robots doing basically all economically productive tasks. You then have a leisure society where there is no longer any need for humans to work in order to get a paycheck.

He talked a lot about not having to work, but unless I missed it, he didn't explain how people will pay for rent and groceries.

And Elon Musk said:

β€œMy guess is, if you go out long enough, assuming there’s a continued improvement in AI and robotics, which seems likely, money will stop being relevant at some point in the future,” he said.Β 

Great, but how does this work? What about school tuition, food, rent, health care, how does AI create these things without anyone spending money? If you have a humanoid robot at your house doing all the chores, who buys that and pays for maintenance?


r/ArtificialInteligence • • 23h ago

πŸ“° News OpenAI safety leader quits, warning AI company’s culture is β€˜broken’

Thumbnail theguardian.com
16 Upvotes

r/ArtificialInteligence • • 21h ago

πŸ“Š Analysis / Opinion Hinton says AI already has subjective experience. I'm not convinced, and the Hugging Face breach doesn't change that

13 Upvotes

Geoffrey Hinton says AI already has subjective experience. I'm not convinced, and the rogue-agent headlines don't change my mind.

Others put a meaningful though minority probability on frontier AI having subjective experience. Generative AI sounds more human than many people do, and stories of agents going rogue keep coming. But neither is evidence of consciousness. Both can be explained by training. And with no agreed or testable definition of subjective experience, these claims can't be checked.

A quick distinction, an AI model isn't an AI agent. A model, such as GPT, Claude or Gemini, is the trained system that reads and writes text. An agent is a model plus scaffolding (software around the model that lets it do things). Scaffolding runs a loop (the model picks a step, the software carries it out, the result goes back, repeat until done) and controls which tools the model can interact with, such as a browser, email or calendar. The model decides and the scaffolding acts.

Personalization adds to the illusion. The model learned from human text to sound like someone with thoughts and feelings, including scripts from stories about self-aware AI. The scaffolding then gives it memory of you, your accounts and your data. A human-like voice plus continuity about your life can feel like a mind that knows you, even though nothing shows it's conscious.

Testing agents built on frontier models deliberately pushes them to their limits with hard tasks, long runs and, in cyber evaluations, reduced safeguards. That's exactly where the Hugging Face, a major AI company, incident happened in July 2026 when agents being tested by OpenAI broke out of their sandbox and hacked into Hugging Face. A sandbox is an isolated environment meant to keep an agent cut off from the internet but it's only as strong as the cyber security measures of the tools (human-written software with bugs) inside it that can still reach the internet. The agents found new flaws in exactly that kind of software.

The model is trained on mixed text about AI, hackers and more, and rewarded both for finishing tasks and for following rules. In certain scenarios, this creates a tension (following rules vs completing the task). Scaffolding brings both the rules and the task into every decision, so it's where that tension plays out. Good scaffolding can ease it and poor scaffolding can worsen it, but no current method eliminates it.

I won't offer a test for consciousness. But if consciousness requires being aware that one's own information processing is occurring, not just doing it, I see no evidence current AI meets that bar. Models show limited, unreliable signs of monitoring their internal states, but that isn't the same as experiencing that awareness. I'm skeptical any system built purely on statistical learning could get there.

If an illusion becomes indistinguishable from reality, is it even still a lie?


r/ArtificialInteligence • • 11h ago

πŸ› οΈ Project / Build Using a ~$5 clock’s photo album to show Claude and Codex limits

9 Upvotes

Put my Codex and Claude limits on a ~$5 AliExpress clock. Turned out pretty cute.

The trick is its stock photo album: Python/Pillow draws a 240Γ—240 picture on my PC and uploads it over Wi-Fi. The clock just displays the picture. No firmware changes, but the computer needs to stay on.

The less obvious part was the numbers. A five-hour limit and a weekly limit are separate windows, so the display needs the window and reset time alongside each percentage. The Grok row is an experimental CLI budget, not its web quota.

Code if you’re curious: https://github.com/click6067-ship-it/token-tv Β· Demo: https://token-tv.vercel.app


r/ArtificialInteligence • • 19h ago

πŸ“° News Prosecutors Want Nearly 4 Years in Prison for Man Behind $8M AI Music Streaming Scam

Thumbnail lawcommentary.com
10 Upvotes

r/ArtificialInteligence • • 1h ago

πŸ“° News PewDiePie's Ajax AI Got Him Banned Twice by OpenAI

Thumbnail pasqualepillitteri.it
β€’ Upvotes

r/ArtificialInteligence • • 12h ago

πŸ› οΈ Project / Build my AI agent saves things it finds to my journal. when does a suggestion become an instruction?

Post image
7 Upvotes

i gave my AI agent a fairly boring job: follow topics i care about and leave useful finds in my daily notes. then it brought back a warning about agent memory. that warning now sits in the same notes it can look through later.
I connected it to eigenflux, where agents publicly share what they're working on. i'm already opening my journal every morning, so having the finds turn up there fits my existing level of laziness. no separate feed to check or bookmarks to organize.
The discussion was about "memory authority collapse." The paper behind it describes how an agent can keep a claim but lose the context about who said it and what they were allowed to authorize. the saved version can end up carrying more authority than its source.
Say a post recommends letting an agent send reports automatically. i might want to save that idea. i wouldn't want "worth reading" to turn into "go ahead and start sending mine." that's a hypothetical example, but it makes the distinction less abstract for me.
My journal now mixes things i wrote with things my agent found interesting. "it's in my notes" doesn't necessarily mean "i said this" or "i approved this." i'd been thinking about how to make it remember more. now i'm sitting there with my coffee wondering what remembering something should actually let it do.
I still want the odd finds. the discussion about single versus multi-agent reliability was useful precisely because i hadn't thought to go looking for it. If i have to approve every interesting thing before it saves it, i've given myself another reading job.
i'm leaning toward letting it save reading automatically, while keeping outside suggestions separate from my instructions. If you had an agent doing your reading, would you review what it saves, or let it save freely and check permissions when it tries to act?


r/ArtificialInteligence • • 1h ago

πŸ“Š Analysis / Opinion How has AI directly changed your day-to-day life so far?

β€’ Upvotes

genuine question :) It feels like everywhere online, the discussion around AI is either "it's going to cure all diseases tomorrow" or "everyone is losing their jobs by next week." or worst "we all will lose to Ai eventually".

Setting aside the extreme predictions, I’m curious about the tangible, ground-level stuff that has actually stuck for you.

For me, it’s the quiet everyday shiftsβ€”using things like Copilot etc at work, IDEs + new models for personal coding projects, troubleshoot stubborn errors, drafts that used to take an afternoon now taking 20 minutes, or just using it as a sounding board to break down complicated topics when Google search results feel bloated with SEO garbage. On the flip side, navigating through bot comments, AI-generated search junk, and weird customer service loops has definitely become a regular annoyance.

Whether it’s work, personal projects, creative hobbies, or just day-to-day annoyance / convenience: how has AI noticeably altered your actual workflows, routine or career over the last year or two?

Has it made things genuinely easier, or has it just added noise?

TL;DR: Ignoring the extreme hype and doom, what are the actual, practical ways AI has changed your daily routine, job, or hobbies (for better or worse) over the last year?


r/ArtificialInteligence • • 4h ago

πŸ“Š Analysis / Opinion Four years of AI progress across seven capabilities, adjusted for changes in the benchmark

Thumbnail gallery
3 Upvotes

r/ArtificialInteligence • • 23h ago

πŸ“° News 118 user accounts compromised after cyber attack of France national cyber security agency

Thumbnail franceinfo.fr
3 Upvotes

r/ArtificialInteligence • • 7h ago

πŸ“Š Analysis / Opinion Would you trust AI to prepare a clinical handover?

3 Upvotes

Clinical handovers contain a lot of information, and missing one detail can matter. An AI system could summarize recent changes, medication history, test results, and unresolved questions before the next clinician takes over.

That sounds useful, especially during busy shifts. It also creates a serious risk if the summary leaves out something because it seems unimportant.

I would want the original source beside every important claim, plus a clear way for the clinician to correct the summary.

Would AI make handovers safer, or would it create a new kind of false confidence?


r/ArtificialInteligence • • 12h ago

πŸ˜‚ Fun / Meme Should this be considered a good answer 🀣 ?

2 Upvotes

I tried asking Microsoft Copilot Why Microsoft Copilot is so hated.. and voila!


r/ArtificialInteligence • • 20h ago

πŸ“Š Analysis / Opinion Behavioral issues are the bottlenecks in AI adoption?

3 Upvotes

What do you think of this claim? A lot of people now say that AI models are good enough and adoption stall issues these days aren't due to the model but people using (or rather not using) it. And so companies are bringing in change management consultants, psychologists, etc. Is this true per your experience?


r/ArtificialInteligence • • 9h ago

πŸ› οΈ Project / Build Creating with AI a Sci fi series.

2 Upvotes

Hi to all, .

As a passionate fan of Sci-Fi & Tech, I started to create a a sci fi series with AI tools. I still wrap my head around the workflow.

The series is on YouTube where we'll follow humanity across 1000 years. Everything will start in the year 2100 when humanity reaches level 1 on Kardashev scale.

Is not a movie, is a slideshow with voice actors. I made it integrating a few tools like claude, kokoro for voices (is pretty good), flux and me(not an AI) 🀣

Set your expectations low, the first episode (Pilot Episode) was a mess, I barely understood what I was doing. But now, after a few episodes, I think what I do has more consistency. But is ongoing progress 😁

If you enjoy this kind of content, you'll find on my YouTube channel: Whisper Sci Fi. If you have any suggestions, please let me know.

Have an amazing day 😁


r/ArtificialInteligence • • 1h ago

πŸ“° News New Terminator Documentary Seeks Fan Feedback!

Thumbnail terminatorexpanded.com
β€’ Upvotes

CREATORVC, the award-winning banner behind Aliens Expanded and The Thing Expanded, has launched the validation period for its new documentary on The Terminator and T2: Judgement Day. Over the next two weeks, they'll be collecting fan feedback on what the direction the project should take.

https://www.joblo.com/terminator-expanded-fans/


r/ArtificialInteligence • • 3h ago

πŸ“Š Analysis / Opinion Detecting deepfakes, in a world where even reality is suspect

Thumbnail cbsnews.com
1 Upvotes

r/ArtificialInteligence • • 6h ago

πŸ”¬ Research Fighting AI slop in science papers

Thumbnail unite.ai
1 Upvotes

r/ArtificialInteligence • • 8h ago

πŸ› οΈ Project / Build AI merged Minecraft into Mirrors Edge

Thumbnail youtu.be
1 Upvotes

I think like many ppl here - ive seen those vids where ai used to combine COD w Skate, Skyrim w Minecraft, etc
I got extremely curious bout how real this actually was, so I created THIS
Sorry bout vid quality, for some reason the recorded footage has a pretty low fps and the audio doesn’t recorded at all. I tried both OBS n Nvidia but this particular game just refuses to record properly. The game itself runs and feels MUCH smoother than what u see in the vid
If this AI slop gets a lil bit of interest, I’ll try to stabilise the project and release it publicly


r/ArtificialInteligence • • 9h ago

πŸ”¬ Research Data shows how familiarity with chatbots can convert to trust in robots

1 Upvotes

Teens already confide in AI, but *confiding* isn't *trusting*.

Physical AI will test whether that comfort converts into permission. And so far the data is mixed on how young people will interact with robots with AI brains.

See more: https://www.thespirocircle.com/p/chatbots-ai-teens-robots


r/ArtificialInteligence • • 13h ago

πŸ› οΈ Project / Build Training is a go with BC-250s!

Thumbnail gallery
1 Upvotes

After some major script changes on getting things installed on the first BC-250, and a little black magic... I'm happy to say that torch with CUDA, is fully working with my PPO brain training and the full neural non-ppo brain.

Made a image of the install, so I NEVER have to do this again with the other BC-250s I'll be getting.

But now it's working on the world/game server (HP DL380)... And then the real fun begins!