Then we went to 100B parameters on globally distributed A100s.
Then we went to 100B parameters on globally distributed A100s. Orion-100B trained at about 65% of the speed of a centralised cluster, at roughly a third of the cost. For a lot of teams, that trade decides whether a model gets built at all.

More from the feed
Taosis indexes announcements, articles, posts and releases about Bittensor subnets from their own channels and keeps a record of each. The text above is the source’s; the figures are read from the chain by Taosis.