Join the conversation

Join the community of Machine Learners and AI enthusiasts.

Sign Up
sequelbox 
posted an update 5 days ago
Post
2944
NEW RELEASES for the new Muse Glimmer 30B!

- Esper 4, our flagship agentic coder: specialist in coding, architecture, DevOps, and MLOps!
- Tachibana-Agent, trained only on code for dedicated, predictable deployment!

GET OUR NEW MODELS:
ValiantLabs/Muse-Glimmer-30B-Esper4
sequelbox/Muse-Glimmer-30B-Tachibana-Agent

Get the datasets for your own training:
sequelbox/Titanium4-DeepSeek-V4-Pro
sequelbox/Mitakihara2-DeepSeek-V4-Pro
sequelbox/Tachibana4-DeepSeek-V4-Pro

We'll be expanding Esper 4 to more models and releasing new models as funding allows - donate for more, faster, better models and datasets: sequelbox/SupportOpenSource

Also, starting work on the SV4 datasets :)

More to come soon!

love,
allegra

Interesting release direction. For a fair comparison, I’d love to see the agentic coding results broken out by task type, context length, tool-call success rate, and recovery from failed runs—not only headline model quality. Those details help teams decide whether a local coder is ready for production.

·

this is a great suggestion for evaluation, thank you!

in general, detailed evaluation is an allocation question for us - with very limited funds currently available for compute, our general priority is to put all of them into datasets and models, to maximize the impact we can make on the open source ecosystem surviving and eventually winning. that is the whole point of our work, to achieve this result :) we can certainly provide more context by offering more benchmarks etc, but ultimately the user will always be better positioned to evaluate their own use cases with their own frameworks and areas of expertise. will still do it if we can!

also, qwen 3.8 releases will be coming soon! literally no money right now 😅 as soon as we can!