โ‰ˆ the.bay.news

99% token accuracy, zero learning. Field notes from fine-tuning vision models with RL.

DEV Community
99% token accuracy, zero learning. Field notes from fine-tuning vision models with RL.
Over the past year I have been fine-tuning open vision-language models - 9B dense up to a 35B...

0 comments

Sign in to join the discussion โ€” your thebay.events account works here.

No comments yet.