castform now supports supervised finetuning (sft)
our goal’s to give developers everything necessary to post-train models for their tasks. rl is all the rage these days (& castform makes it super easy to use), but sft's one of most compute-efficient ways to teach a model exactly
the post-training platform for the ai engineer
Joined June 2025
- how we're spending time improving onboarding for our devs 👇
- teaching castie to roll a die with reinforcement learning:lots of talk about agi, asi, rsi but ask any frontier LLM to roll a die and it will almost always say "4." claude, gpt, kimi - doesn't matter, 4.4.4.4. so here's how i post-trained a model to reliably roll a die (i.e. each number ~1/6th of the time) & why it's a nice sandbox for
- woohoo free credits to finetune your own custom model :)
- castform is in open beta! our goal’s to enable any developer post-train their own llms. in the world of rapidly rising llm costs & providers guarding capabilities, we believe the ability to shape model behavior shouldn’t be a privilege. this release today is our small step



