Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

There isn't a fast churn in the underlying pretrained model, nor RL. It's mostly orchestration around the model. Said another way you could just pretrain and RL for longer.

Also I believe there is both a market for extremely fast local inference with current model performance and that such fast inference would unlock unforeseen usecases. Especially as TPS approaches early computer clock cycles and data rates.



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: