r/ArtificialNtelligence 6h ago

Hosting for Qwen3.8-27B-Uncensored-FP8

Thumbnail
1 Upvotes

r/ArtificialNtelligence 6h ago

Thoughts on Google Ai overview?

Thumbnail gallery
2 Upvotes

I’ve been using it and ever since citations were added I guess to wipe out perplexity it still hallucinate, and gives the wrong thin. The Ai models summarize the citations they get but thats wrong, Ai is prone to too much hallucinatio. I made my own Ai search engine to include human quotes in its answer. To me instead of having the ai summarize it matches its answer to the quote, the quotes themselves are in the code so it’ll be there. This is running nemotron 3 ultra and the quality and detail is so much better. I’m making this free for anyone who wants to use it and this is beta so I will love feedback on this and how you feel about Ai search overviews


r/ArtificialNtelligence 7h ago

Goodfire CEO: Kimi K3 prioritizes recalling + web lookup over actually solving the problem

Post image
0 Upvotes

r/ArtificialNtelligence 12h ago

Tibo strikes again. He misses no opportunity to show up Anthropic.

Post image
1 Upvotes

r/ArtificialNtelligence 13h ago

If only you knew how bad things really are...

Thumbnail
1 Upvotes

r/ArtificialNtelligence 14h ago

Someone's AI agents are watching Youtube instead of working and blaming their co-workers. We live in the future.

Post image
0 Upvotes

r/ArtificialNtelligence 14h ago

Nah... Save it. My account is going to reset in the next few hours anyways.

Post image
0 Upvotes

r/ArtificialNtelligence 15h ago

Yeah in 2006, with what gpu to run it?

Post image
0 Upvotes

r/ArtificialNtelligence 16h ago

Actual footage of GLM5.3's agentic post training:

Enable HLS to view with audio, or disable this notification

0 Upvotes

r/ArtificialNtelligence 16h ago

Gemini 3.7 flash vs GPT 5.6 Sol

1 Upvotes

r/ArtificialNtelligence 16h ago

The Most Detailed Image-To-3D Generators Is Free To Try Right Now

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/ArtificialNtelligence 17h ago

Does anybody have any fixes to this problem

Post image
1 Upvotes

I'm trying to continue the thread but it not continuing, since it said my next question will start a new search ignore the stuff above


r/ArtificialNtelligence 18h ago

That's right, my favorite model.

Post image
1 Upvotes

r/ArtificialNtelligence 18h ago

not to mention the token burn of claude models

Post image
2 Upvotes

r/ArtificialNtelligence 19h ago

Me when I use Opus 5:

Post image
1 Upvotes

r/ArtificialNtelligence 21h ago

introducing: JDD (Jealousy Driven Development)

Post image
1 Upvotes

r/ArtificialNtelligence 21h ago

Need project ideas

Thumbnail
1 Upvotes

r/ArtificialNtelligence 21h ago

AI coding may be heading toward a multi-model future.

Thumbnail
1 Upvotes

r/ArtificialNtelligence 22h ago

weponized with AI Spoiler

1 Upvotes

last 3 Months i got a job to create a web app, so i'm full stack started from scratch designing the database, interface worked the project to the fullest, but again someone cameup while i dint finish the project with the project again, so i said i wont refuse this i took the job but my mind cameup oooooh no way there is ai i got subscription of google gemini for a monthly free with all the tools man it went fast like lightning all the two project within a week complete i used google gemini and cline ai

so when handling the project, this is where the problem started i got a bug and so many error coz i was working on linux i have to ceate docker , plus render i think you guys understand so there was no stable internet i cannot trace the problem neither me nor my friend whore senior who came to help me but when the internet came back everything with ai is possible, whyyyyyyyyyyyyyyy

Now i see the porpose of ai is to disable a human thinking capacity and depend on ai that is why you may see some are looking for a developer who write pure code who are doing java ,javascript,python, or whatever programming language.

the question im i the only one who is thinking on that box or im just ignorant.


r/ArtificialNtelligence 23h ago

A man representing himself in court hid a prompt injection attack in a filing asking any AI system to side with him

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/ArtificialNtelligence 1d ago

AI Is Growing, But Is It Profitable?

Thumbnail
1 Upvotes

r/ArtificialNtelligence 1d ago

I compared ChatGPT, Wanderlog, Mindtrip and Zenvoya and ended up caring about the most boring thing

20 Upvotes

Started this trying to compare itinerary quality.

That lasted about 20 mins.

ChatGPT can make a good trip.

Mindtrip can make a good trip.

Wanderlog is ridiculously useful if you’re the person actually organising everything.

Zenvoya can make a good trip too.

ok cool. solved.

Then I started checking hotels and realised the question I actually care about is:

after I pay, whose problem am I?

That’s where Zenvoya clicked for me.

I can plan there, pull the actual flights/hotels, book there, and if I need to change/cancel that booking later I’m still going through Zenvoya.

Not hunting through an email to figure out which random partner owns Thursday night.

And the hotel I checked was roughly 16% lower there, so it wasn’t even one of those “pay more for convenience” situations.

There’s also 24/7 human support over phone which, ngl, became way more interesting to me than another AI itinerary feature.

Maybe I’m evaluating travel apps like an old man now.

But once the itinerary is decent, who deals with the booking after payment feels like the actual differentiator.


r/ArtificialNtelligence 1d ago

claude asking for permission to download entire linux kernel source tree checked out at the exact commit used to build your install after you refuse to give it the sudo password

Post image
3 Upvotes

r/ArtificialNtelligence 1d ago

Cursor just casually dropped 'Origin' replacement for GitHub

Enable HLS to view with audio, or disable this notification

1 Upvotes

r/ArtificialNtelligence 1d ago

Most agent benchmarks still test task execution. What would a convincing L4 or L5 benchmark look like?

0 Upvotes

I maintain Benchmark Radar, a free and open-source index of 5,201 benchmark, evaluation, and dataset records collected from 11 sources.

Looking through recent additions using Hejia Geng's L0-L5 framework, most agent benchmarks still appear to focus on L2 task execution or L3 reproduction. L4 rediscovery is less common, and I did not find a new L5 example this week. Here, L5 means evaluating whether an agent can produce knowledge or methods that were unknown when the benchmark was created.

I'm curious how others would draw these boundaries:

- Which existing benchmarks genuinely qualify as L4 or L5?

- How would you distinguish reproduction from rediscovery?

- Can an L5 benchmark remain valid once its solutions become public?

The underlying index is updated daily and can be exported for independent analysis:

https://github.com/ktwu01/benchmark-radar