-
Q-Transformer: Scalable Offline Reinforcement Learning via Autoregressive Q-Functions
Paper • 2309.10150 • Published • 24 -
In-Context Pretraining: Language Modeling Beyond Document Boundaries
Paper • 2310.10638 • Published • 28 -
Farzi Data: Autoregressive Data Distillation
Paper • 2310.09983 • Published • 7 -
LLaVA-Plus: Learning to Use Tools for Creating Multimodal Agents
Paper • 2311.05437 • Published • 45
Mat Miller
matdmiller
AI & ML interests
None yet
Organizations
Collections
1
spaces
4
models
None public yet
datasets
None public yet