Blog
LLMs & Texto
Variable Bit-width Quantization: Learning Per-Group Precision for "Bigger-but-Smaller" Language Models
arXiv:2607.02893v1 Announce Type: new Abstract: Low-bit quantization shrinks language models but treats precision as a single global hyper-parameter: every weight uses the same bit-width. We introduce Variable Bit-width Quantization (VBQ), a training-time method in which each contiguous group of 64 weights learns its own resolution from {1,2,4,8} bits via a Gumbel-Softmax relaxation, trained jointly by an alternating optimization that gives the precision logits a clean, task-aligned signal. VBQ ...
arXiv cs.LG
·Hamish Ogilvy
·
// relacionados
Leia também
Blog
As alegações mais escandalosas no processo da Apple contra a OpenAI por segredos comerciais
Blog
O que a mais recente descoberta em IA da Anthropic mostra — e o que não mostra
Blog
Novo guia de prompts da OpenAI diz aos usuários para parar de complicar e começar pelo resultado
Blog