depthwise-convolution

Tag

Cards List
#depthwise-convolution

Convolution for Large Language Models

arXiv cs.CL · 2026-07-22 Cached

This technical report explores adding lightweight depthwise convolution to the query/key/value projections in Transformer blocks for LLMs, providing local inductive bias that improves downstream accuracy with negligible parameter cost.

0 favorites 0 likes
← Back to home

Submit Feedback