Academic

Academic

Academic · 약 1분

AE-LLM: Adaptive Efficiency Optimization for Large Language Models

arXiv:2603.20492v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved remarkable success across diverse applications, yet their deployment remains challenging due to substantial computational …

Kaito Tanaka, Masato Ito, Yuji Nishimura, Keisuke Matsuda, Aya Nakayama
조회수 40회
Academic · 약 1분

Distributed Gradient Clustering: Convergence and the Effect of Initialization

arXiv:2603.20507v1 Announce Type: new Abstract: We study the effects of center initialization on the performance of a family of distributed gradient-based clustering algorithms introduced in …

Aleksandar Armacki, Himkant Sharma, Dragana Bajovi\'c, Du\v{s}an Jakoveti\'c, Mrityunjoy Chakraborty, Soummya Kar
조회수 35회
Academic · 약 1분

Delightful Distributed Policy Gradient

arXiv:2603.20521v1 Announce Type: new Abstract: Distributed reinforcement learning trains on data from stale, buggy, or mismatched actors, producing actions with high surprisal (negative log-probability) under …

Ian Osband
조회수 56회
Academic · 약 1분

Does This Gradient Spark Joy?

arXiv:2603.20526v1 Announce Type: new Abstract: Policy gradient computes a backward pass for every sample, even though the backward pass is expensive and most samples carry …

Ian Osband
조회수 37회
Academic · 약 1분

Towards Practical Multimodal Hospital Outbreak Detection

arXiv:2603.20536v1 Announce Type: new Abstract: Rapid identification of outbreaks in hospitals is essential for controlling pathogens with epidemic potential. Although whole genome sequencing (WGS) remains …

Chang Liu, Jieshi Chen, Alexander J. Sundermann, Kathleen Shutt, Marissa P. Griffith, Lora Lee Pless, Lee H. Harrison, Artur W. Dubrawski
조회수 33회