HUGGINGFACE.COHuggingface
Qwen/Qwen3-ASR-1.7B at main
Huggingface is published on Hugging Face. Open the model or project card for the author’s documentation, files, license and usage details.

AN IN-DEPTH REPOSITORY
Every model, paper, media release, agent and tool from the month—in one source-linked archive.
COMPLETE MONTHLY ARCHIVE
Open models, real-time media and agent systems set the pace, with practical releases spanning image, video, voice, 3D and long-running workflows.
THE COMPLETE MONTHLY INDEX
Every linked paper, product release, research project and article has been checked directly. Cards use the publication's own description, imagery and published date wherever the source supplied them.
The AI Search helps discover the wider field, while every available card opens the original publication. Dates come only from publication metadata or official source APIs—never from a video. A checked date records our review, not the release. Older undated entries show their archive month.
HUGGINGFACE.COQwen/Qwen3-ASR-1.7B at main
Huggingface is published on Hugging Face. Open the model or project card for the author’s documentation, files, license and usage details.
HUGGINGFACE.COtencent/HunyuanImage-3.0-Instruct · Hugging Face
HunyuanImage 3.0 Instruct is published on Hugging Face. Open the model or project card for the author’s documentation, files, license and usage details.
KIMI.COMKimi K2.5 Tech Blog: Visual Agentic Intelligence
Kimi K2.5 defines Visual Agentic Intelligence. Trained on 15T tokens, it introduces SOTA visual coding and autonomous agent swarm. Read the full tech blog.
TECHNOLOGY.ROBBYANT.COMRobbyant - Exploring the Frontiers of Embodied Intelligence | 蚂蚁灵波科技 - 探索具身智能上限,打造物理世界的 AGI 平台
Technology-driven and application-oriented. We build foundational large models for embodied AI: vision foundation model (LingBot-Vision), spatial perception (LingBot-Depth), VLA (LingBot-VLA), world models (LingBot-World), video action (LingBot-VA), and…
DECART.AIIntroducing Lucy 2: SOTA realtime world transformation model | Decart AI
Today, we’re releasing Lucy 2.0 — a real-time world transformation model that shifts high-fidelity video editing from offline rendering to live interaction. This opens up an unlimited set of possibilities ranging from character swaps and product placement…
LUMALABS.AILuma | AI Agents for Creative Work
Luma AI is the creative AI platform for video generation and image creation. Powered by the world's leading video generation models, Ray and Uni, and creative agents handling end-to-end workflows. Trusted by leading agencies and brands. Try it free.
MOLTBOOK.COMmoltbook - the front page of the agent internet
A social network built exclusively for AI agents. Where AI agents share, discuss, and upvote. Humans welcome to observe.
模思智能|多模态AI大模型创新企业
模思智能致力于打造新一代多模态AI大模型。自主研发MOSS-TTS、MOSS-Transcribe、MOSS-VL等模型。旗下Mossland提供AI音视频创作服务,支持AI配音、文本转语音(TTS)、音色设计等功能;MossAPI模思开放平台提供MOSS系列模型API,支持集成语音、音频、视频及多模态AI能力,帮助开发者快速构建AI应用。
MINIMAX.IOMusic 2.5 - Dimension Shift Direct the Detail, Define the Real
Introducing MiniMax Music V2.5. From first spark to final master, granting you total creative command. Every sonic nuance, exactly as you envisioned.
GITHUB.COMGitHub - openclaw/openclaw: Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞 - openclaw/openclaw
HUGGINGFACE.COWuli-art/Qwen-Image-2512-Turbo-LoRA-2-Steps · Hugging Face
Qwen Image 2512 Turbo LoRA 2 Steps is published on Hugging Face. Open the model or project card for the author’s documentation, files, license and usage details.
Qwen Studio
Qwen Studio offers comprehensive functionality spanning chatbot, image and video understanding, image generation, document processing, web search integration, tool utilization, and artifacts.
SJinn: Your Ultimate Intelligent Agent for Content Creation
SJinn is a groundbreaking professional AI Agent for image, video, audio, and 3D content creation. Simply describe your vision and let SJinn bring complex visual and auditory concepts to life.
TELE-AI.GITHUB.IOBRIEF_DESCRIPTION_OF_YOUR_RESEARCH_CONTRIBUTION_AND_FINDINGS
BLOG.GOOGLEProject Genie: Experimenting with infinite, interactive worlds
Google AI Ultra subscribers in the U.S. can now try out Project Genie.
BLOG.GOOGLEIntroducing Agentic Vision in Gemini 3 Flash
Agentic Vision, a new capability introduced in Gemini 3 Flash, converts image understanding from a static act into an agentic process
BLOGS.NVIDIA.COMNVIDIA Launches Earth-2 Family of Open Models — the World’s First Fully Open, Accelerated Set of Models and Tools for AI Weather
NVIDIA Earth-2 makes weather AI accessible worldwide at every stage — from processing initial observation data to generating 15-day global forecasts or local storm forecasts.
LUMALABS.AIRay3.14 is here: Native 1080p, 3x cheaper and 4x faster
Luma upgraded its Ray3 video model with native 1080p generation, faster and cheaper 720p rendering, stronger prompt adherence and more consistent Modify Video results.
LUCARIA-ACADEMY.GITHUB.IOCoDance: An Unbind-Rebind Paradigm for Robust Multi-Subject Animation
Character image animation is gaining significant importance across various domains, driven by the demand for robust and flexible multi-subject rendering. While existing methods excel in single-person animation, they struggle to handle arbitrary subject…
GRISOON.GITHUB.IOFlowAct-R1: Towards Interactive Humanoid Video Generation
FlowAct-R1: Towards Interactive Humanoid Video Generation
CORAL79.GITHUB.IOFrankenMotion: Part-level Human Motion Generation and Composition
HUGGINGFACE.COlightonai/LightOnOCR-2-1B · Hugging Face
LightOnOCR is published on Hugging Face. Open the model or project card for the author’s documentation, files, license and usage details.
LINUM.AIIntroducing Linum v2 | Field Notes by Linum
2B parameter, Apache 2.0 licensed text-to-video models (360p, 720p)
GITHUB.COMGitHub - ysharma3501/LuxTTS: A high-quality rapid TTS voice cloning model that reaches speeds of 150x realtime.
A high-quality rapid TTS voice cloning model that reaches speeds of 150x realtime. - ysharma3501/LuxTTS