Zum Hauptinhalt springen

Showing 1–4 of 4 results for author: Winsor, E

Searching in archive cs. Search in all archives.
.
  1. arXiv:2312.10091  [pdf, other

    cs.IR cs.CL cs.LG

    Look Before You Leap: A Universal Emergent Decomposition of Retrieval Tasks in Language Models

    Authors: Alexandre Variengien, Eric Winsor

    Abstract: When solving challenging problems, language models (LMs) are able to identify relevant information from long and complicated contexts. To study how LMs solve retrieval tasks in diverse situations, we introduce ORION, a collection of structured retrieval tasks spanning six domains, from text understanding to coding. Each task in ORION can be represented abstractly by a request (e.g. a question) tha… ▽ More

    Submitted 13 December, 2023; originally announced December 2023.

  2. arXiv:2211.12312  [pdf, other

    cs.LG cs.AI

    Interpreting Neural Networks through the Polytope Lens

    Authors: Sid Black, Lee Sharkey, Leo Grinsztajn, Eric Winsor, Dan Braun, Jacob Merizian, Kip Parker, Carlos Ramón Guevara, Beren Millidge, Gabriel Alfour, Connor Leahy

    Abstract: Mechanistic interpretability aims to explain what a neural network has learned at a nuts-and-bolts level. What are the fundamental primitives of neural network representations? Previous mechanistic descriptions have used individual neurons or their linear combinations to understand the representations a network has learned. But there are clues that neurons and their linear combinations are not the… ▽ More

    Submitted 22 November, 2022; originally announced November 2022.

    Comments: 22/11/22 initial upload

  3. arXiv:2110.15343  [pdf, other

    cs.LG

    Scatterbrain: Unifying Sparse and Low-rank Attention Approximation

    Authors: Beidi Chen, Tri Dao, Eric Winsor, Zhao Song, Atri Rudra, Christopher Ré

    Abstract: Recent advances in efficient Transformers have exploited either the sparsity or low-rank properties of attention matrices to reduce the computational and memory bottlenecks of modeling long sequences. However, it is still challenging to balance the trade-off between model quality and efficiency to perform a one-size-fits-all approximation for different tasks. To better understand this trade-off, w… ▽ More

    Submitted 28 October, 2021; originally announced October 2021.

    Comments: NeurIPS 2021

  4. arXiv:1905.04746  [pdf, other

    math.CO cs.DM

    Generalized Lyndon Factorizations of Infinite Words

    Authors: Amanda Burcroff, Eric Winsor

    Abstract: A generalized lexicographic order on words is a lexicographic order where the total order of the alphabet depends on the position of the comparison. A generalized Lyndon word is a finite word which is strictly smallest among its class of rotations with respect to a generalized lexicographic order. This notion can be extended to infinite words: an infinite generalized Lyndon word is an infinite wor… ▽ More

    Submitted 20 June, 2019; v1 submitted 12 May, 2019; originally announced May 2019.

    Comments: 14 pages, 1 figure

    MSC Class: 05A05; 68R15