Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
samsam55 's Collections
Retrieval
Image Evaluation
Benchmarks
Long Offline Context Understanding
Personas
Accessibility
Games
Finance/Trading Agents
Lower Requirements Inference Engine & Model
Video Models Capability Surveys
Long Horizon Agent Memory Harnesses & Techniques
Streaming Video Understanding
Cyber
VLM (image+text => text)
OCR
Image
Small but smart (?) models
Text to Music
Skills
Video Generation & Pipelines
Coding Agents (Games)
Reinforcement Learning Etc..
Datasets
Self Improving
Run on CPU Optimizations
Deep Search
World View Creation (out painting 3D)
Computer Use
Coding LLMs
Visual Multi Modal LLM
TTS & Speech to Text
Misc
Agents
3D Models & Modeling

Video Generation & Pipelines

updated 4 days ago
Upvote
-

  • Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration

    Paper • 2605.17423 • Published May 17 • 30

  • Video Generation Models: A Survey of Post-Training and Alignment

    Paper • 2610.00812 • Published 11 days ago • 61

  • In-Distribution Forcing for Long Video Generation at Test Time

    Paper • 2610.03120 • Published 9 days ago • 49

  • APRIL-AIGC/T3-Video

    Text-to-Video • Updated 6 days ago • 207 • 23
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs