Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

NousResearch

Team
company
https://nousresearch.com/
nousresearch
Activity Feed

AI & ML interests

We are dedicated to advancing the field of natural language processing, in collaboration with the open-source community, through bleeding-edge research and a commitment to symbiotic development.

Papers

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation

Targeted Neuron Modulation via Contrastive Pair Search

View all Papers

karan4d's profile picture Teknium's profile picture Shannon Sands's profile picture Jeffrey Quesnelle's profile picture Jai 's profile picture Bowen Peng's profile picture Nobody.png's profile picture dmayhem93's profile picture dillon rolnick's profile picture Ben's profile picture Jon Durbin's profile picture Robin Fernandes's profile picture Sam Herring's profile picture yoni's profile picture théo gigant's profile picture neuralink's profile picture Morgane Moss's profile picture Siddharth's profile picture snav's profile picture Rohan's profile picture 𒐪's profile picture
NousResearch 's papers 7
Submitted by
théo gigant
11

Decoupling the Benefits of Subword Tokenization for Language Model Training via Byte-level Simulation

NousResearch NousResearch
2
Submitted by
Jeffrey Quesnelle
17

Targeted Neuron Modulation via Contrastive Pair Search

NousResearch NousResearch
33 2
Submitted by
Bowen Peng
48

Efficient Pre-Training with Token Superposition

NousResearch NousResearch
8
Submitted by
Bowen Peng
31

Long Context Pre-Training with Lighthouse Attention

NousResearch NousResearch
65 2
Submitted by
Sumuk Shashidhar
55

Hermes 4 Technical Report

NousResearch NousResearch
3
Submitted by
AK
60

Hermes 3 Technical Report

NousResearch NousResearch
8
Submitted by
AK
85

YaRN: Efficient Context Window Extension of Large Language Models

NousResearch NousResearch
518 4
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs