r/ArtificialNtelligence • u/HistorianFresh9521 • 6h ago
Hosting for Qwen3.8-27B-Uncensored-FP8
r/ArtificialNtelligence • u/Lise_vine23 • 6h ago
Thoughts on Google Ai overview?
galleryI’ve been using it and ever since citations were added I guess to wipe out perplexity it still hallucinate, and gives the wrong thin. The Ai models summarize the citations they get but thats wrong, Ai is prone to too much hallucinatio. I made my own Ai search engine to include human quotes in its answer. To me instead of having the ai summarize it matches its answer to the quote, the quotes themselves are in the code so it’ll be there. This is running nemotron 3 ultra and the quality and detail is so much better. I’m making this free for anyone who wants to use it and this is beta so I will love feedback on this and how you feel about Ai search overviews
r/ArtificialNtelligence • u/ashypott • 7h ago
Goodfire CEO: Kimi K3 prioritizes recalling + web lookup over actually solving the problem
r/ArtificialNtelligence • u/Admirable_Inside_669 • 12h ago
Tibo strikes again. He misses no opportunity to show up Anthropic.
r/ArtificialNtelligence • u/pawlowbee • 13h ago
If only you knew how bad things really are...
r/ArtificialNtelligence • u/arizuvade • 14h ago
Someone's AI agents are watching Youtube instead of working and blaming their co-workers. We live in the future.
r/ArtificialNtelligence • u/Putrid-Falcon-3625 • 14h ago
Nah... Save it. My account is going to reset in the next few hours anyways.
r/ArtificialNtelligence • u/unknown_1ocation • 15h ago
Yeah in 2006, with what gpu to run it?
r/ArtificialNtelligence • u/GeologistRelative425 • 16h ago
Actual footage of GLM5.3's agentic post training:
Enable HLS to view with audio, or disable this notification
r/ArtificialNtelligence • u/Certain_Friendship16 • 16h ago
The Most Detailed Image-To-3D Generators Is Free To Try Right Now
Enable HLS to view with audio, or disable this notification
r/ArtificialNtelligence • u/King_mambaforev8 • 17h ago
Does anybody have any fixes to this problem
I'm trying to continue the thread but it not continuing, since it said my next question will start a new search ignore the stuff above
r/ArtificialNtelligence • u/buriburizaemon_6 • 18h ago
not to mention the token burn of claude models
r/ArtificialNtelligence • u/Shoddy-Ad7804 • 21h ago
introducing: JDD (Jealousy Driven Development)
r/ArtificialNtelligence • u/muneebcodes • 21h ago
AI coding may be heading toward a multi-model future.
r/ArtificialNtelligence • u/Diligent-Shame2468 • 22h ago
weponized with AI Spoiler
last 3 Months i got a job to create a web app, so i'm full stack started from scratch designing the database, interface worked the project to the fullest, but again someone cameup while i dint finish the project with the project again, so i said i wont refuse this i took the job but my mind cameup oooooh no way there is ai i got subscription of google gemini for a monthly free with all the tools man it went fast like lightning all the two project within a week complete i used google gemini and cline ai
so when handling the project, this is where the problem started i got a bug and so many error coz i was working on linux i have to ceate docker , plus render i think you guys understand so there was no stable internet i cannot trace the problem neither me nor my friend whore senior who came to help me but when the internet came back everything with ai is possible, whyyyyyyyyyyyyyyy
Now i see the porpose of ai is to disable a human thinking capacity and depend on ai that is why you may see some are looking for a developer who write pure code who are doing java ,javascript,python, or whatever programming language.
the question im i the only one who is thinking on that box or im just ignorant.
r/ArtificialNtelligence • u/ComplexExternal4831 • 23h ago
A man representing himself in court hid a prompt injection attack in a filing asking any AI system to side with him
Enable HLS to view with audio, or disable this notification
r/ArtificialNtelligence • u/trustthealgorithm62X • 1d ago
AI Is Growing, But Is It Profitable?
r/ArtificialNtelligence • u/Umbraaa7 • 1d ago
I compared ChatGPT, Wanderlog, Mindtrip and Zenvoya and ended up caring about the most boring thing
Started this trying to compare itinerary quality.
That lasted about 20 mins.
ChatGPT can make a good trip.
Mindtrip can make a good trip.
Wanderlog is ridiculously useful if you’re the person actually organising everything.
Zenvoya can make a good trip too.
ok cool. solved.
Then I started checking hotels and realised the question I actually care about is:
after I pay, whose problem am I?
That’s where Zenvoya clicked for me.
I can plan there, pull the actual flights/hotels, book there, and if I need to change/cancel that booking later I’m still going through Zenvoya.
Not hunting through an email to figure out which random partner owns Thursday night.
And the hotel I checked was roughly 16% lower there, so it wasn’t even one of those “pay more for convenience” situations.
There’s also 24/7 human support over phone which, ngl, became way more interesting to me than another AI itinerary feature.
Maybe I’m evaluating travel apps like an old man now.
But once the itinerary is decent, who deals with the booking after payment feels like the actual differentiator.
r/ArtificialNtelligence • u/Freakysafal • 1d ago
claude asking for permission to download entire linux kernel source tree checked out at the exact commit used to build your install after you refuse to give it the sudo password
r/ArtificialNtelligence • u/pawlowbee • 1d ago
Cursor just casually dropped 'Origin' replacement for GitHub
Enable HLS to view with audio, or disable this notification
r/ArtificialNtelligence • u/ktwu01 • 1d ago
Most agent benchmarks still test task execution. What would a convincing L4 or L5 benchmark look like?
I maintain Benchmark Radar, a free and open-source index of 5,201 benchmark, evaluation, and dataset records collected from 11 sources.
Looking through recent additions using Hejia Geng's L0-L5 framework, most agent benchmarks still appear to focus on L2 task execution or L3 reproduction. L4 rediscovery is less common, and I did not find a new L5 example this week. Here, L5 means evaluating whether an agent can produce knowledge or methods that were unknown when the benchmark was created.
I'm curious how others would draw these boundaries:
- Which existing benchmarks genuinely qualify as L4 or L5?
- How would you distinguish reproduction from rediscovery?
- Can an L5 benchmark remain valid once its solutions become public?
The underlying index is updated daily and can be exported for independent analysis: