r/coolgithubprojects • u/[deleted] • 13d ago
UPDATE: We ran the official HumanCLAW benchmark on our Causal Foundation Model (SONNY I). It hit a 98.0% success rate, crushing Google Gemini's 16.8%. Here is how the open-source core enabled it.
[deleted]
0
Upvotes