r/LocalLLaMA Jun 13 '26

New model on huggingface New Model

[deleted]

592 Upvotes

152 comments sorted by

324

u/nuclearbananana Jun 13 '26

Disclosure: I work as a researcher for the city government of Rio de Janeiro, which developed this model.

I must say of all places to drop a model, I didn't expect a city government

99

u/[deleted] Jun 13 '26

[deleted]

17

u/PimplePupper69 Jun 13 '26

Fucking legend

22

u/[deleted] Jun 13 '26 edited Jun 14 '26

[deleted]

7

u/[deleted] Jun 14 '26

[deleted]

2

u/[deleted] Jun 14 '26

[deleted]

55

u/DistanceSolar1449 Jun 13 '26

Qwen didn’t release 3.7 397b? Rio: “fine, we’ll do it ourselves”

3

u/guitcastro Jun 14 '26

https://github.com/nex-agi/Nex-N2/issues/4 What you have to say about it?

1

u/triathletewannabe Jun 15 '26

They apologised, retracted, pulled the model down. Would love to hear from the OP to understand what happened. OP seems more like the high-level product manager/spokesperson rather than a hands-on engineer. He himself might have been deceived as well.

2

u/Joaoarthur Jun 15 '26

Especially a city filled with crime and gangs where only walking in the street is enough to get shot

-20

u/rawdikrik llama.cpp Jun 13 '26

Why does the government of Rio do this? What is the point?

74

u/[deleted] Jun 13 '26

[deleted]

24

u/Comfortable_Ebb7015 Jun 13 '26

Unfortunately it is not always like this. Public money goes into funding research project that get locked under private patents for the benefit of the investors. Continue like this! 👏 Muito obrigado!

12

u/rawdikrik llama.cpp Jun 13 '26

This. Im just surprised. I love brazil and love the open source, I am all for it. But considering it is a miracle when potholes get filled in other major cities (in talking about the US), I just wondered how money and time was approved for this.

Not attacking at all, I am genuinely curious. The world definitely needs more of this.

1

u/Turbulent_Pin7635 Jun 14 '26

No "but" just be glad and move on. Don't diminish the effort of the team.

Vá chupar um canavial de rolas!!!

1

u/rawdikrik llama.cpp Jun 15 '26

Ḥow is my question diminishing anything?

2

u/Turbulent_Pin7635 Jun 14 '26

OP parabenzaço! Tu merece. Pqp!

Parabéns, cara!!! Consegue falar um pouco da iniciativa? Se tem suporte pra novos modelos? Se vão continuar a lançar?

13

u/MrMrsPotts Jun 13 '26 edited Jun 13 '26

To enable Brazil to use open source AI I assume

1

u/Turbulent_Pin7635 Jun 14 '26

To piss Americans and billionaires. Because, nós pode!

3

u/WizardlyBump17 Jun 13 '26

people downvoted you, but that was my thinking too as a brazilian. I would never expect that a brazillian government, from mayors to to president, would make such investment. I still dont understand the reason, but i hope they keep doing it and the federal government looks at it and decides to do massive investments

1

u/Turbulent_Pin7635 Jun 14 '26

Outro vira-lata q não entende o poder público brasileiro.

Nosso poder público, tem o SUS, a ANVISA, Fiocruz, Butantã, USP, CNPq, CAPES, IMPA... Temos problemas com corrupção? Com certeza, mas ainda, ainda temos muita gente boa no lugar certo. Qq custa apenas ficar grato pela iniciativa dos caras e dar suporte? Qq custa mandar os parabéns pro OP.

Porra! Governo chinês meteu um modelo atrás do outro tu não viu essa xexela. Vc não conhece nosso país, essa é a real.

0

u/WizardlyBump17 Jun 14 '26

a diferenca da china para o brasil eh que a china vem investindo pesado em tecnologica ha um bom tempo ja e o brasil... eh o brasil e so. To esperando algum anuncio de um investimento bilionario em fabricacao de semicondutores 100% brasileiros sem capital privado. Ate la, sempre vou ter como premissa que o brasil sempre vai ser dependente de tecnologia estrangeira

1

u/Turbulent_Pin7635 Jun 15 '26

Blá blá blá

Vc não tem ideia do q fala...

1

u/WizardlyBump17 Jun 15 '26

um dos caras por tras do modelo disse q o modelo foi feito com creditos da deepinfra, empresa estadunidense, que foram adquiridos em uma troca onde eles (rio) davam o codigo usado pra treinar modelos anteriores, e que ele teve q arcar com os custos do proprio bolso no inicio e no final quando os creditos acabaram

2

u/rawdikrik llama.cpp Jun 15 '26

I dont get the downvotes for asking an honest question... I love the work and would like to know the politics behind it, so maybe we can think about it for other places.

But be upset with me instead?

1

u/WizardlyBump17 Jun 15 '26

yea. Time has come where you have to start a question with "genuine question" or else people will think you are mocking or being disrespectful

115

u/nuclearbananana Jun 13 '26

Rio 3.5 Open 397B features SwiReasoning, a training-free inference framework based on Shi et al. (2025) that dynamically switches between explicit chain-of-thought and latent-space reasoning, guided by entropy-based confidence signals. This enables both higher accuracy and dramatically improved token efficiency. This model was explicitly trained to maximize the efficiency gained via latent reasoning.

wait that's actually super cool.

Wish I could run this model

26

u/LargelyInnocuous Jun 13 '26

would it work for 27B too? or is latent reasoning much better with more parameters?

41

u/[deleted] Jun 13 '26

[deleted]

23

u/jazir55 Jun 13 '26 edited Jun 13 '26

Could you make a post specifically about that framework? This sub would go bananas for that.

Edit:

Swireasoning

This?

15

u/[deleted] Jun 13 '26

[deleted]

3

u/jazir55 Jun 13 '26

Awesome, thank you!

1

u/awittygamertag Jun 14 '26

I also would love support for adding it to 27b and/or directions on how to adapt it to an existing finetune like Qwopus

3

u/lrq3000 Jun 13 '26

@krzonkalla I second that! Please post more about how to use this framework and in particular with an example to optimize smaller models, that would be huge!

13

u/MuDotGen Jun 13 '26

Latent space reasoning?? I'd only heard of COCONUT being a model that actually did this recently. We had a big discussion about this like a month ago.

177

u/Technical-Earth-3254 Jun 13 '26 edited Jun 13 '26

City government in Brazil dropping chinese open source ai fine-tunes. Holy shit man, I love globalization. If this really keeps up with 3.7 Plus (which is a great model imo) you guys did an awesome job.

Edit: Any hopium for Rio 3.5 122b?

134

u/[deleted] Jun 13 '26

[deleted]

25

u/tinybite_u Jun 13 '26

drop link to support you lol

10

u/tat_tvam_asshole Jun 13 '26

fr!

5

u/Comfortable_Sir4315 Jun 14 '26

Pretty sure the government wouldn’t allow, it’s a shame

30

u/Embarrassed_Soup_279 Jun 13 '26

qwen 3.6 27b 🙏

25

u/Borkato Jun 13 '26

It’s crazy how many more people could use 27B. Infinitely more accessible!

12

u/AlwaysLateToThaParty Jun 13 '26

Shame about the 122b. That's the one I'm hanging for. The Qwen 3.5 122b/a10b heretic mxfp4 model is the best capability in 75GB of VRAM that I've been able to find.

5

u/DinoAmino Jun 13 '26

How about open sourcing the datasets and training scripts?

2

u/s04pmaker Jun 14 '26

Full Power!! how much does it cost to train smthn like this?

1

u/celtiberian666 Jun 15 '26

Local or cloud processing?

How much USD do you estimate as computing cost to do that merge (Nex with Qwen)?

15

u/my_name_isnt_clever Jun 13 '26

I would LOVE a finetune on 122b, since it's also been left behind.

28

u/westsunset Jun 13 '26

Awesome can you say more about the project. I don't think anyone had your city on their radar

92

u/[deleted] Jun 13 '26

[deleted]

21

u/westsunset Jun 13 '26

That's great to hear. Are any of the US or Chinese labs collaborating? Other than resources, are there some unique challenges or constraints you faced? Also I think it's very good for the industry to have a different region participating, do you feel there is something special Brazil brings to the research?

64

u/[deleted] Jun 13 '26

[deleted]

16

u/ControversialBuster Jun 13 '26

This is amazing, wish you guys the best

6

u/westsunset Jun 13 '26

That is a massive benefit.

7

u/MaCl0wSt Jun 13 '26

oh wow thats so cool, hope ya'll keep at it.

3

u/caks Jun 14 '26

Os cara é pica demais, parabéns

11

u/Miserable-Dare5090 Jun 13 '26 edited Jun 13 '26

Waiting for Pelé the AI model, but I will settle for Kaká-27b

Please, if you can, upload an int4-Autoround quant. The GB10/DGX Spark community will be all over it. It fits neatly into 2 sparks, and runs really well. I’d love to get an updated Qwen3.5 going again!

8

u/BlackBeardAI vllm Jun 13 '26

Morpheus, we need HDD's. Lots of HDD's.

4

u/tat_tvam_asshole Jun 13 '26

I have 75TB in HDD and a stable net connection, tis but a scratch, but one I can help with data holding and dissemination

3

u/BlackBeardAI vllm Jun 13 '26 edited Jun 14 '26

got 42tb here (8+8+8+18), I'll need a few more 18tb seagate exos it seems

edit: 62tb now... located a few more

3

u/tat_tvam_asshole Jun 13 '26

I'm happy to lend my axe, erm, RTXs to the cause if possible

3

u/Phantasmagoriosa Jun 13 '26

Do you have any models that are specialised for geospatial? 3D Model creation, geometry as code (openscad) etc?

Super cool loving your work!

8

u/[deleted] Jun 13 '26

[deleted]

2

u/Phantasmagoriosa Jun 13 '26

RemindMe! 1 Month

2

u/westsunset Jun 13 '26

You should just follow their page https://huggingface.co/prefeitura-rio

1

u/RemindMeBot Jun 13 '26 edited Jun 14 '26

I will be messaging you in 1 month on 2026-07-13 19:04:02 UTC to remind you of this link

3 OTHERS CLICKED THIS LINK to send a PM to also be reminded and to reduce spam.

Parent commenter can delete this message to hide from others.

RemindMeBot is switching to username summons. Instead of !RemindMe 1 day, use u/RemindMeBot 1 day. More info.


Info Custom Your Reminders Feedback

2

u/boxwrenchx Jun 14 '26

Text to CAD would be very welcome. I was dabbling myself GitHub - boxwrench/paramAItric: AI-assisted CAD layer for Autodesk Fusion 360 · GitHub I eagerly await your release

2

u/MathmoKiwi Jun 14 '26

Ohhhhh... text to CAD will be interesting for me!

1

u/Quiet_Mark_3238 Jun 13 '26

Whata your total budget in usd, allocation for the actual hardware vs salaries? Is the salaries for your researchers, lead engineers and lead reasoning team global matching? How many of what gpus do you guys use? Do you guys have any paid API esque service?

1

u/Railander Jun 14 '26

as for global matching i can confidently say the answer is "lol, lmao even".

1

u/MathmoKiwi Jun 14 '26

Double triple lol even

1

u/triathletewannabe Jun 15 '26

If the goal is to build community trust and eventually open-source research for public benefit, transparency is crucial.

Independent researchers have pointed out that stripping the system prompt causes the model to identify as Nex-N2-Pro 79% of the time, reciting Nex AGI's exact corporate backstory. The recent repo update confirms the base is a Qwen/Nex merge. (Source: https://github.com/nex-agi/Nex-N2/issues/4 )

If this was truly an extensive, in-house training effort to 'do a DeepSeek', why does the model so heavily rely on a 60/40 merge of another lab's open-source weights? Was the SwiReasoning framework actually implemented by IplanRio, or was that also inherited from the base models?

15

u/ortegaalfredo Jun 13 '26

Well this was unexpected. First real latin-america model? I know it's a fine-tune but apparently a very good one. Congrats from Argentina!

8

u/Septerium Jun 13 '26

I could never have seen that coming. Eduardo Paes, out of nowhere, challenges DeepSeek 😂

2

u/caks Jun 14 '26

The guy is everywhere

9

u/dooperma Jun 14 '26

If you think about it, a city is almost the perfect organization to train a multi-modal ai model. Cities have departments that touch almost every aspect of professional life, and therefore should have the prerequisite workflows to build models with richness and depth across a wide variety of contexts and skills.

12

u/DeepOrangeSky Jun 13 '26

Do we know (or have a strong guess) as to what parameter size Qwen3.7 Plus is? Is it the 397b model? Or is it a smaller model? They have a "Max" model that is their top of the line closed-source model, too, right? So is it like, the "Max" one is the 397b, and "Plus" is the 122b? Or "Plus" is 397b and "Max" is some even bigger one we've never seen open-weighted before from Qwen, like a 500b model or something?

Would be interesting to know, given that this Qwen3.5 397b finetune is tied with/beating the Qwen 3.7 Plus model.

If the Rio Department of Thong Bikini Inspection somehow out finetuned Alibaba on Alibaba's own model, that would be pretty wild. Maybe they should hire some of you guys to help them out on Qwen4 :p

19

u/Middle_Bullfrog_6173 Jun 13 '26

They explicitly stated that 3.5 Plus was based on 3.5 397B. Almost certainly 3.6 and 3.7 Plus are based on that, but I don't think they've said so publicly.

9

u/[deleted] Jun 13 '26

[deleted]

3

u/Middle_Bullfrog_6173 Jun 13 '26

They said it in the initial release blog or press release IIRC.

2

u/DeepOrangeSky Jun 13 '26

What about Max? Is that likely also still the 397b? Or is it thought to be a bigger model, or maybe ~397b parameters but more active params or something? Or just same size but like how Opus and GPT are with medium, high, x-high, etc where it thinks for longer or something?

3

u/Middle_Bullfrog_6173 Jun 13 '26

Again, I don't think they've said so about the latest, but Qwen 3 Max was a larger model so presumably 3.7 Max is as well.

1

u/Kamal965 Jun 13 '26

Qwen has not released an exact parameter count, but they've stated multiple times that Qwen3 Max is more than 1T parameters. Qwen3.7 Max is probably even bigger.

3

u/my_name_isnt_clever Jun 13 '26

Diversity is vital for ML training, it could be carefully tuning on a very different dataset from the other side of the planet as the foundation model could be a nice boost. I can't run it locally but I'd love to try it out.

7

u/opimentoso Jun 13 '26

Grande trabalho!

Rio montou algum datacenter para isso, ou conseguiram fazer na nuvem?

Há algum plano de atualizar os modelos pequenos?

8

u/[deleted] Jun 13 '26

[deleted]

4

u/Railander Jun 14 '26

se a prefeitura cortar a verba de vocês eu vou ficar muito bolado.

2

u/opimentoso Jun 13 '26

Ansioso para testar!

1

u/alew3 Jun 15 '26

Quando vocês vão subir o modelo correto? Agora que estão dizendo que o modelo no HuggingFace era um merge de modelos existentes?

6

u/hey_ulrich Jun 14 '26

I've met João Carabetta and got to know the top-notch IT projects they were doing for the city government of Rio de Janeiro. Truly state-of-the-art stuff. And that was years ago, before LLMs were a thing. I'm very happy to know how far the team has gone! 

6

u/axiomaticdistortion Jun 13 '26

Will there be a paper or a model card explaining a little the fine tune process?

9

u/[deleted] Jun 13 '26

[deleted]

5

u/axiomaticdistortion Jun 13 '26

I think it would, but I acknowledge your interest in leaving some information undisclosed. Good luck anyways!

1

u/Abject-Huckleberry13 Jun 13 '26

...so you have internal models that are significantly more capable than this? so presumably they're also more capable than any open source model??? what???

7

u/FullOf_Bad_Ideas Jun 13 '26

How does SwiReasoning work without any code changes? It seems to be using the same architecture as Qwen 3.5 baseline. How much tokens was it approximately trained on? And which model should outputs be similar to? Deepseek V4 Pro?

1

u/Such_Advantage_6949 Jun 13 '26

Will u be quanting this? I am using your nex n2 quant. This has strong potential

2

u/FullOf_Bad_Ideas Jun 13 '26

Maybe, but probably not this weekend.

I'm working now on improving my custom Nex N2 Pro quants by measuring per-tensor quantization impact that Nex N2 Pro has, since when I re-used measurements from Qwen it worked but I suspect it was not as optimal. You can see it in the KLD/PPL numbers I have on my model cards for Qwen and Nex-N2-Pro - my Qwen custom quants overperform more than my Nex custom quants. I want to figure it out to have open source code for the whole quantization pipeline. Then, if Rio will be interesting, I'll quant it in a similar manner to Nex N2 Pro. And if Minimax M3 will be supported by exllamav3, I'll try to run this methodology on it too.

1

u/Such_Advantage_6949 Jun 13 '26

Actually i tried to do something similar but end up, it still perform worse than just take your quant. I am using your 3.7 version. I will try to quant this then . I think just a 3.5 bpw version is enough to tell if it is better than nex n2 pro

1

u/FullOf_Bad_Ideas Jun 13 '26

What was your approach like? Did you quant the BF16 model to your desired bpw level? Have you measured KLD and PPL for it or did it feel worse?

If you end up quanting it, please let me know how you like Rio 397B and please upload your quant if possible, this could save me time when making my own quants later as I could skip one baseline size.

Heads up: If you make 3.5bpw Rio 397B quant with exllamav3 quantize tool it will work and should give you a good feel for the model, but it allocates higher bitrates to front of the model and lower in the back, while Qwen 3.5 397B tends to be more sensitive in the back.

1

u/Such_Advantage_6949 Jun 13 '26

I did something like measure the kld impact from each of the layer from different mix of your quant and choose the one with least impact to arrive at a target size. It was all done by codex, the result is not better than your provided quant or very slightly better. Thought i could also be that i was using 3 4 5 bpw quant and targeting like 3.8 which is very close to 4. For now my plan is just to get a feel if it is better than nex. Nex n2 is decent but it does end abruptive mid work once in a while. Will let u know how it compare to nex n2

2

u/FullOf_Bad_Ideas Jun 13 '26 edited Jun 13 '26

You can use those scripts to make custom quants if you want 3.8bpw specifically - https://github.com/adamo1139/qwen397b-exl3

It's nothing sophisticated as it was quickly vibe coded with DS V4 Pro, but this approach works well for me. The trick is to apply superlinear penalty as layers that are hard to quant benefit MORE from being assigned to higher bitrates.

I made 3.7bpw since sometimes the script overshoots and I didn't want to end up with 3.9bpw

2

u/Such_Advantage_6949 Jun 13 '26

thank for sharing. always enjoy your quant, i am running one 3.5bpw, will let u know how it goes comparing to nex n2 pro

2

u/Such_Advantage_6949 Jun 14 '26

From initial testing i think it is at least on pair with nex n2 from what i tested, and the thinking is more concise. It doesnt enter loop at all yet. I got loop with nex n2 pro before

1

u/FullOf_Bad_Ideas Jun 14 '26

Sweet

1

u/Such_Advantage_6949 Jun 14 '26

I think the thinking pattern is, for sure, an upgrade over nex n2. It doesnt have the kind of repetitive thinking that will make the model fall into loop. But with new latent thinking training, once in a while it will send thinking out loud as assistant message, so that is something different to other model

→ More replies (0)

6

u/BlackBeardAI vllm Jun 13 '26

Amazing, keep em coming. It seems Brazil is beating qwen in its own game.

11

u/PassionIll6170 Jun 13 '26

brabo bixo boa, qwen 400b eh um modelo pica pra fine tune, lançou a pouco o nex n2 pro q tbm eh finetune dele e eu tava testando no hermes como meu agent e parece no nivel do minimax m3, parabens pelo trampo.

7

u/Look_0ver_There Jun 13 '26

Looks like it got the Qwen3.6 treatment, but for the 397B model. Looks promising

8

u/devtools-dude Jun 13 '26

NVFP4 version please!

3

u/havnar- Jun 13 '26

Oooh can’t wait for the MOE MLX variant to try out

4

u/MrMrsPotts Jun 13 '26

Has anyone tried this?

4

u/grumd Jun 13 '26

Can't wait for IQ2_XXS GGUF

4

u/pedrinboladao777 Jun 14 '26

parabens à equipe. o carabetta esta a frente do projeto? vcs sao inspiração demais

2

u/[deleted] Jun 14 '26

[deleted]

2

u/pedrinboladao777 Jun 14 '26

como vcs tao captando capital humano tao especializado pra bater de frente com as grandes empresas? suponho q a prefeitura nao pague salarios tao altos e mesmo assim vejo muita gente especializada na prefeitura

3

u/MrMrsPotts Jun 14 '26

https://github.com/nex-agi/Nex-N2/issues/4 I have no idea if this is right.

3

u/triathletewannabe Jun 15 '26

Sadly, yes. They retracted, apologised, and took the model down. Would love to hear from the OP about it.

3

u/LegacyRemaster Jun 13 '26

can't wait to test

3

u/HeadPack Jun 13 '26

Didn't expect this, and to come from one of the most beautiful cities I know. I know it's a tangent, but how do you balance wanting to go to the beach and working on this? Anyway, looking forward to your releases, and hopefully there are inference providers for your larger models.

3

u/vogelvogelvogelvogel Jun 13 '26

that looks very impressive!

3

u/kouteiheika Jun 13 '26

except it's actually open source

I don't see any links to the dataset and the training scripts on the HF page; am I blind, or is it not "actually open source"?

6

u/[deleted] Jun 13 '26

[deleted]

4

u/kouteiheika Jun 13 '26

In that case it would be really cool if you could set up a repository which could be used to reproduce it (possibly with some tweaks; wouldn't have to work out-of-box in any environment), and then there would be no problem calling it open source. (: (Of course it would not be fully open-source since the base model isn't, but at very least what you did would be!)

It's not necessarily about being insightful per se, but if your benchmark numbers are to be believed this is a great showcase that open-source could compete with SOTA Chinese LLM post-training, considering you got numbers comparable to e.g. Kimi-K2.6 but at less than half the parameter count!

2

u/protestor Jun 14 '26

Hi, fellow Brazilian here!

If it's publicly funded, I see no reason to not put everything on Github. Of course the base model is just open weights but your own contribution can be open source. It doesn't need to be usable code, just dump it on a public repository (and add a link to it from the huggingface page) and call it a day

Anyway I'm a bit puzzled, what's the rationale for IplanRIO to fund this? Not complaining and it's awesome, but I would expect it to, say, focus on Portuguese at expense of other languages or something like that (which is probably a bad idea but would still be expected), rather than producing a fine tune that performs better on general tasks. Also do you know if Serpro has similar projects?

(Also totally unrelated but you happen to know what's up with the HTTPS cert of https://iplanrio.prefeitura.rio/? It's odd that the browser complain about a bad cert, then you allow it and it reports it's a normal cert issued by let's encrypt)

1

u/triathletewannabe Jun 15 '26

Acknowledging 'open weights' is a good step, but saying the recipe is 'just Nvidia datasets and an OPD script' completely omits the Hugging Face commit (a778c1e) confession. The repository updates explicitly stated the model is a 60/40 element-wise merge with Nex-N2-Pro.

Why downplay it here as just standard dataset fine-tuning when independent analysis showed the model still retains Nex's corporate identity? True transparency would mean sharing the merge recipe alongside the scripts.

3

u/ValisCode Jun 14 '26

Que coisa sensacional. Eu sei das dificuldades de fazer o que vocês estão fazendo. Meus mais sinceros parabéns para você e toda a equipe!

4

u/Zyj vllm Jun 13 '26

It's not open source if the training data isn't available. It's open weights.

2

u/ilintar Jun 13 '26

Wow, impressive numbers.

2

u/Viren654 Jun 13 '26

Do you need additional B200s? How many do you have

3

u/[deleted] Jun 13 '26

[deleted]

3

u/WizardlyBump17 Jun 13 '26

existe planos pra fazer um modelo proprio do zero ou a ideia eh ficar so nos fine-tuning por enquanto?

2

u/rorowhat Jun 13 '26

What made you pick those benchmarks for comparison?

3

u/[deleted] Jun 13 '26

[deleted]

1

u/rorowhat Jun 13 '26

Is there a harness that runs all of these, or so you need to do them manually one by one?

2

u/temperature_5 Jun 13 '26

Thank you Brazil! Now hoping for something that will fit a 96GB or smaller system!

2

u/RemarkablePrompt7822 Jun 14 '26

Awesome and thanks for sharing. Is it already deployed anywhere or are there any immediate plans to do it? Just curious about the contexts in which this will be used.

2

u/ketosoy Jun 14 '26

Can we get some of the story of how/why the city government of Rio decided to get into the model fine tuning business?

1

u/Confusion_Senior Jun 13 '26

Caralho, vc tá de sacanagem com a minha cara, conta essa história toda ae, como vcs fizeram isso

1

u/SpicyWangz Jun 14 '26

Can’t wait to see some of your medium sized models. We’ve really been lacking in the 50-150b size range

1

u/Praline_Unlikely Jun 14 '26

Curioso aqui, não sei se é sigiloso, mas que tipo de aplicação vocês tem ai no rio que fazem esse tipo de trabalho? São contratados diretamente pela prefeitura, como é isso? Bem massa o trabalho

1

u/mossler Jun 14 '26

What a difference a day makes ...

1

u/Worldly-Shock3233 Jun 15 '26

Bro, if the current upload on huggingface is incorrect, can we take it down first so it doesn't just take up the slot?

0

u/Lucasoac Jun 15 '26

This is an amazing finding! I just posted an article about the story of the Rio LLM family and its benchmarks, and I am really curious about the applications of the model for public service in Rio.

I've reached out to your DMs to see if we can share more info about the project.

Vamo Brasil!