AI & ML interests

JAX, Flax, TPU, 🤗

Recent Activity

flax-community's activity

odellus 
posted an update 14 days ago
view post
Post
1514
Tired: shitposting on bsky
Wired: shitposting on hf
  • 1 reply
·
christopher 
posted an update 2 months ago
view post
Post
1656
The folks at Foursquare released a dataset of 104.5 million places of interest ( foursquare/fsq-os-places) and here's all of them on a plot
·
christopher 
posted an update 2 months ago
christopher 
posted an update 5 months ago
view post
Post
1327
4 million chess puzzles
morgan 
posted an update 7 months ago
view post
Post
1303
Llama 3.1 405B Instruct beats GPT-4o on MixEval-Hard

Just ran MixEval for 405B, Sonnet-3.5 and 4o, with 405B landing right between the other two at 66.19

The GPT-4o result of 64.7 replicated locally but Sonnet-3.5 actually scored 70.25/69.45 in my replications 🤔 Still well ahead of the other 2 though.

Sammple of 1 of the eval calls here: https://wandb.ai/morgan/MixEval/weave/calls/07b05ae2-2ef5-4525-98a6-c59963b76fe1

Quick auto-logging tracing for openai-compatible clients and many more here: https://wandb.github.io/weave/quickstart/