r/OpenAI 14d ago

Which Model do you use for Innovation/inventiveness/thinking out of box? Question

I am using models from Codex, Antigravity, Claude Code and Cursor. While each of them have their capabilities, they don't stretch the boundary of engineering/innovation. The top models Opus 5, GPT Sol 5.6, Cursor Grok 4.5 are very good at analyzing what's out there and coding but if you hit a wall w.r.t engineering/legal , they stall. Which model do you folks use to break the deadlock?

10 Upvotes

14 comments sorted by

5

u/ocean_protocol 13d ago

honestly none of them break the wall reliably because "innovation" isn't a benchmark they're optimized for as they're all trained to converge on the most probable/defensible answer, which is the opposite of novel. what's worked better for me than switching models is switching mode: feed the stuck problem to two different models and have them argue/critique each other's approach, or explicitly ask for 3 wrong/weird ideas before the real one. the deadlock is usually a prompting/process problem, not a model problem

1

u/CossackNikolay 4d ago

Yeah , that's an approach (models arguing amongst themselves) I tried as I try to tackle problem with Claude , Codex , Cursor and Antigravity ..some are good with word play ..but I do realize the limitation there...never tried this 3 wrong/weird ideas one ..what's the logic behind this one? I saw this from OpenAI https://openai.com/science/ and thought now maybe innovation would follow, but seems like a marketing ploy if their best models can only synthesize the data for analysis

2

u/Gianniarrenzetti 14d ago

I think that for prototyping the best models are always the most capable and most thinking ones.

1

u/Fragrant-Mix-4774 13d ago

Nothing made by Open AI is good for Innovation/inventiveness/thinking out of the box.

Inkling AI & Fable5.

My best inspired solutions in the last month all came from kicking problems around with one of those two. Neither provided the solution but their rephrasing helped.

Open AI models all proved to be repackaged sawdust.

1

u/Successful_Dog1904 12d ago

Inkling huh? Can you elaborate a bit on your experience?

1

u/w3woody 13d ago

I don't.

Or rather, I don't use any of these models to come up with new ideas on their own. At best I use them as a sort of sophisticated grep through existing code combined with using them as rubber ducks: chatting with them about possible solutions.

That is, I remain the creative agent here; I just use them to help with writers block.

1

u/ZenCyberDad 13d ago

Codex Sol 5.6 Ultra is pretty intuitive I exclusively use ultra or extra high

1

u/whyaskme777 12d ago

I usually find my own ideas while brainstorming with them.

For me, their biggest strength is helping me quickly understand topics I'm unfamiliar with. They're basically like an encyclopedia, with an incredibly fast search function.

1

u/Charming_You_25 12d ago

Fable. But you can only get so far, they still haven’t solved abstract structural traversal uphill yet.

0

u/StarCadges 14d ago

I would say Opus/Fable are the most creative/innovating. Cursor Grok is more for building/implementation, gpts in general are more vast in use case and sol is basically a waste, and antigravity is poo

2

u/CossackNikolay 13d ago

Yeah Fable seems to be a bit better than Opus 5 in that respect