Please turn JavaScript on

The Berkeley Artificial Intelligence Research Blog

Is this your feed? Claim it!

Publisher:  Unclaimed!
Message frequency:  0.1 / day

Message History


Figure 1: CUDA-to-MLX optimization translation map. CUDA optimization knowledge can be translated into architecture-native MLX strategies rather than copied instruction-for-instruction.

We face a new epoch in computing. Hardware is changing rapidly — not just faster GPUs, but a growing range of chips from different vendor...


Read full story

Overview of ABBEL compared to traditional recursive summarization. Beliefs replace the full interaction history as the agent’s working context, and belief grading improves performance by supervising the contents of each belief state..

As task horizons grow, LLM contexts can’t scale forever. Self-summarization enables concise, interpretable ...


Read full story