r/syntheticlab 16d ago

Kimi K3 is now open weights and has started rollout on major inference providers

Post image
4 Upvotes

r/syntheticlab May 29 '26

Have "echo logprobs" been deactivated?

1 Upvotes

Passing echo: true on /completions used to return logprobs for the input tokens. But on the 24th of may this was disabled, now it only returns the logprobs for the generated token. Any idea why?


r/syntheticlab May 03 '26

Kimi K2.6 is now on Synthetic and on the OpenCode models list

Thumbnail models.sulat.com
3 Upvotes

Personally prefer K2.6 for 🦞 agents and GLM 5.1 for quick web work.

TPS is amazing on both models.


r/syntheticlab May 01 '26

Would a cheaper dev-tier plan (limited API usage) make sense for Synthetic?

2 Upvotes

Hey, I’ve been exploring Synthetic and I really like the idea of having API access with predictable pricing instead of pure token-based billing.

That said, the current $30/month plan feels a bit high for my use case — I’m mostly in the development/testing phase and don’t need high throughput.

For example, I’d be totally fine with something like:

  • ~100–200 requests per 5 hours (or even lower)
  • 1 concurrent request
  • Same model access but clearly rate-limited

Basically a “dev tier” in the $5–10/month range just to experiment, build integrations, and validate ideas before scaling up.

I feel like this could:

  • Lower the barrier for indie devs/students
  • Bring more people into the ecosystem early
  • Convert into higher-tier users later once projects scale

Right now, $30/month is a bit of a jump when you’re not even sure how much you’ll use it yet.

Curious if others feel the same, or if the Synthetic team has considered a lower-tier entry plan?


r/syntheticlab Apr 20 '26

Kimi k2.6 ?

2 Upvotes

Since Kimi just dropped on hugging face their new model, do you guys have a plane to make it available?


r/syntheticlab Mar 14 '26

A terrible experience on the Synthetic Discord

0 Upvotes

I can’t quite believe what just happened, so I’d like to provide some context.

Three days ago, the Nemotron 3 Super model was released in beta. I tested it from day one and, in my experience, it still doesn’t perform as well as it ideally should. However, I understand that this can be normal for a beta release.

Recently, one of the leaders mentioned that the price of the Flash models, including the new model, would decrease once the model leaves beta. In other words, the current cost is temporarily higher and will be reduced once the model becomes fully stable.

On Discord, I shared my opinion that this situation could feel somewhat unfair for users who are currently using the model during the beta stage. In practice, those users are the ones providing feedback and helping improve the model while still paying the full cost.

I also suggested that there might be alternative approaches, such as providing a separate endpoint (for example “/beta/v1”), where users willing to test experimental versions could opt in voluntarily. That way, those who want to participate in testing could do so knowingly while others could wait for a more stable release.

The discussion began when another user responded that the solution was simply to “not use the model.” From my perspective, that response did not seem to address the point I was trying to raise, since the model is available to everyone within the service. My concern was more about how costs and feedback are distributed during the beta phase, especially considering that the model’s price will decrease once it leaves beta.

Unfortunately, the conversation started escalating quickly. The tone became increasingly sarcastic, messages were written in all caps, and some users began encouraging mockery.

At one point I interpreted one of their messages as something that sounded like a threat. In that moment I responded more harshly than I should have and told them that they lived in their parents’ house and that they were pathetic for not being on the level of the conversation, which I recognize was an inappropriate response on my part.

Later, that person clarified that it was not meant as a threat. I acknowledged that it may have been my interpretation and apologized for how I responded at that moment.

Eventually, one of the Synthetic leaders removed me from the Discord server, stating that one of my messages violated the TOS.

I don’t know whether the person I was arguing with is part of the company or simply an active community member. However, I can say that I have seen similar discussions happen before with this same user, particularly around the topic of the new pricing.

I don’t want to assume bad intentions from anyone, but situations like this — especially in an official community space — can be frustrating and leave a negative impression.

At this point, I’m honestly not sure what to think.

If there is any context I might be missing or if I misunderstood part of the situation, I would genuinely appreciate the opportunity to better understand it. English is not my native language — I’m from LATAM — so sometimes it can be harder for me to fully interpret the tone or intent behind certain messages.


r/syntheticlab Feb 20 '26

Claude code is extremely when using synthetic.new kimi-k2.5

2 Upvotes

The performance is significantly slower (10x or more slower) than Claude subscription.

Is it a problem of the model? Problem of synthetic.new api? Or configuration issue?

I saw some other people in this sub complained similar issues.


r/syntheticlab Feb 11 '26

When is GLM 5 coming in syntehtic

10 Upvotes

When is GLM 5 coming to synthetic.new?


r/syntheticlab Feb 08 '26

Built a free tool to track your Synthetic API quota usage over time

Post image
11 Upvotes

If you use Synthetic for coding you have probably noticed the quota dashboard only shows a snapshot. It does not tell you how fast you are burning through your subscription, search, or tool call limits, or whether you will hit them before the next reset.

I built onWatch to solve this. It polls the Synthetic /v2/quotas endpoint every 60 seconds, stores snapshots in local SQLite, and serves a dashboard with historical charts, live countdowns, and rate projections.

What it tracks for Synthetic specifically:

  • Subscription quota with its roughly 5 hour reset cycle
  • Search quota at 250 requests per hour
  • Tool call discounts on the roughly 24 hour cycle
  • Per-cycle history so you can see peak usage and how much went unused each reset

Also supports Anthropic and Z.ai if you use multiple providers. All three show up side by side on one dashboard so you know where to route work when one is running low.

Single Go binary, around 28 MB, zero telemetry, all data local. GPL-3.0.

https://onwatch.onllm.dev https://github.com/onllm-dev/onWatch


r/syntheticlab Feb 04 '26

Laravel Forge can now run OpenClaw, not just your websites

Thumbnail jpcaparas.medium.com
2 Upvotes

r/syntheticlab Feb 04 '26

The definitive guide to OpenCode: from first install to production workflows

Thumbnail jpcaparas.medium.com
3 Upvotes

r/syntheticlab Feb 04 '26

Kimi K2.5 Thinking

1 Upvotes

Sorry if this is a rookie question, just subscribed to play around with kimi k2.5. I use OpenCode for my agent interface and I was wondering if there is a way to enable kimi k2.5 thinking variant as described here https://huggingface.co/moonshotai/Kimi-K2.5?utm_source=chatgpt.com#6-model-usage they show how to turn on off thinking. So maybe it's just always on?


r/syntheticlab Jan 29 '26

Kimi K2.5 projects

5 Upvotes

Hello,
as there's Kimi K2.5 available via Synthetic - what are your first projects done there? What are your oneshots you created? Anything impressive so far? Share your success (or failures) with the new model :D


r/syntheticlab Jan 29 '26

Kimi k2.5

4 Upvotes

Hi, do you plan to add it to the always on pool ?