Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Centre for Vision, Speech and Signal Processing - University of Surrey

university
https://www.surrey.ac.uk/centre-vision-speech-signal-processing
cvssp_research
Activity Feed Request to join this org

AI & ML interests

Audio, Vision

Recent Activity

anindyamondal  authored a paper about 1 month ago
ABACUS: Adapting Unified Foundation Model for Bridging Image Count Understanding and Generation
sanchit-gandhi  authored a paper 4 months ago
Voxtral TTS
anindyamondal  updated a dataset 4 months ago
cvssp/OmniCount-191
View all activity

Sanchit Gandhi's profile picture haoheliu's profile picture Xinhao Mei's profile picture Yin Cao's profile picture Tavis Shore's profile picture asmar nadeem's profile picture Xubo Liu's profile picture Jinhua Liang's profile picture Anindya Mondal's profile picture Marco Volino's profile picture Xiatian Zhu's profile picture Swapnil's profile picture Peter Bradley's profile picture BimBomBam's profile picture sskela's profile picture Rohan Kalra's profile picture Victor Ricciuti's profile picture kamwoh's profile picture Ahmed Bourouis's profile picture Taran Rai's profile picture Tony Alex's profile picture Frank Lu 呂祖方's profile picture
Organization Card
Community About org cards

CVSSP is an internationally recognised leader in audio-visual machine perception research. With a diverse community of more than 150 researchers, CVSSP is one of the largest audio and vision research groups in the UK.

models 7

cvssp/audioldm2-music

0.3B • Updated Apr 16, 2024 • 811 • 29

cvssp/audioldm2-large

0.7B • Updated Apr 16, 2024 • 1.45k • 20

cvssp/audioldm2

0.3B • Updated Apr 16, 2024 • 20.8k • 71

cvssp/audioldm-l-full

Updated Apr 16, 2024 • 781 • 22

cvssp/audioldm

Updated Apr 16, 2024 • 191 • 34

cvssp/audioldm-s-full-v2

0.2B • Updated Apr 16, 2024 • 2.3k • 21

cvssp/audioldm-m-full

0.4B • Updated Apr 16, 2024 • 500 • 31

datasets 3

cvssp/OmniCount-191

Updated Apr 7 • 34

cvssp/SpaGBOL

Updated Jun 4, 2025 • 13 • 1

cvssp/WavCaps

Viewer • Updated Jul 6, 2023 • 1 • 6.19k • 55
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs