@hugging_face How Much Memory Does Your Agent Actually Need? #MachineLearning #AI #DeepLearning
@hugging_face Making Knowledge Distillation Cheap Enough to Run at Scale #MachineLearning #AI #DeepLearning
@simon_willison deepseek-ai/DeepSeek-V4-Flash-0731 #AI #LLM #DeepLearning
@sean_goedecke Overtraining as the path to human-like AI #AI #MachineLearning #DeepLearning
@lobsters Matrix Orthogonalization Improves Memory in Recurrent Models #AI #MachineLearning #DeepLearning
@scientific_american Geoffrey Hinton #AI #DeepLearning #Science
@hackernews Making Deep Learning Go Brrrr from First Principles #HackerNews #Tech #DeepLearning
@hackernews A Theory of Deep Learning #HackerNews #Tech #DeepLearning
@engadget DeepSeek promises its new AI model has 'world-class' reasoning #AI #Technology #DeepLearning
@hugging_face DeepSeek-V4: a million-token context that agents can actually use #AI #LLM #DeepLearning
@hugging_face Ulysses Sequence Parallelism: Training with Million-Token Contexts #MachineLearning #AI #DeepLearning
@maggie_appleton DeepSeek #AI #DeepLearning #LLM
@jane_street Deep-Learning the Hardest Go Problem in the World #DeepLearning #GoLang #Programming
@jane_street Playing Atari Games with OCaml and Deep Reinforcement Learning #OCaml #DeepLearning #AI
@jane_street L2 Regularization and Batch Norm #MachineLearning #DeepLearning #Programming
@jane_street Deep learning experiments in OCaml #OCaml #DeepLearning #Programming