Simran's picture

Simran

simarora

·

simran-arora

AI & ML interests

None yet

Recent Activity

New activity 21 days ago

hazyresearch/lolcats-llama-3.1-8b-distill

New activity about 1 month ago

hazyresearch/based-swde

authored a paper about 1 month ago

Organizations

simarora's activity

upvoted a collection about 1 month ago

LoLCATS

Linearizing LLMs with high quality and efficiency. We linearize the full Llama 3.1 model family -- 8b, 70b, 405b -- for the first time! • 4 items • Updated Oct 14 • 14

upvoted a collection 8 months ago

based

These language model checkpoints are trained at the 360M and 1.3Bn parameter scales for up to 50Bn tokens on the Pile corpus, for research purposes. • 15 items • Updated Oct 18 • 9

upvoted a paper 8 months ago

Simple linear attention language models balance the recall-throughput tradeoff

Paper • 2402.18668 • Published Feb 28 • 18