Andrej Jovanovic's picture

8

Andrej Jovanovic

maddox-j

·

https://maddox-j.github.io/

maddox-j

AI & ML interests

None yet

Recent Activity

upvoted a paper 8 days ago

DES-LOC: Desynced Low Communication Adaptive Optimizers for Training Foundation Models

upvoted a paper 8 days ago

MT-DAO: Multi-Timescale Distributed Adaptive Optimizers with Local Updates

upvoted a paper 5 months ago

SVD-Free Low-Rank Adaptive Gradient Optimization for Large Language Models

View all activity

Organizations

None yet

upvoted 2 papers 8 days ago

DES-LOC: Desynced Low Communication Adaptive Optimizers for Training Foundation Models

Paper • 2505.22549 • Published May 28 • 1

MT-DAO: Multi-Timescale Distributed Adaptive Optimizers with Local Updates

Paper • 2510.05361 • Published 17 days ago • 1

upvoted 2 papers 5 months ago

SVD-Free Low-Rank Adaptive Gradient Optimization for Large Language Models

Paper • 2505.17967 • Published May 23 • 17

Quartet: Native FP4 Training Can Be Optimal for Large Language Models

Paper • 2505.14669 • Published May 20 • 77

upvoted a paper 7 months ago

Hogwild! Inference: Parallel LLM Generation via Concurrent Attention

Paper • 2504.06261 • Published Apr 8 • 110

upvoted a paper 8 months ago

Panza: A Personalized Text Writing Assistant via Data Playback and Local Fine-Tuning

Paper • 2407.10994 • Published Jun 24, 2024 • 2

upvoted 2 papers 9 months ago

HALO: Hadamard-Assisted Lossless Optimization for Efficient Low-Precision LLM Training and Fine-Tuning

Paper • 2501.02625 • Published Jan 5 • 16

QuEST: Stable Training of LLMs with 1-Bit Weights and Activations

Paper • 2502.05003 • Published Feb 7 • 43