✕ Clear all filters
82 articles
▶ Videos →

📰 ArXiv cs.AI

82 articles · Updated every 3 hours · View all reads

All Articles 189,374Blog Posts 171,258Tech Tutorials 50,780Research Papers 36,772News 23,157 ⚡ AI Lessons
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper 1mo ago
From Generation to Simulation: How Far Are World Models from Being True Simulators?
arXiv:2608.23070v1 Announce Type: new Abstract: With the rapid progress of diffusion models and large-scale video generation, generative world models are increa
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper 1mo ago
MLLM-Guided Semantic Correction for Text-to-Video Generation
arXiv:2608.16513v1 Announce Type: cross Abstract: Recent advances in diffusion models and Transformer architectures have led to significant progress in text-to-
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper 1mo ago
Concept Guidance: Precise, Training-Free Latent Control for Text-to-Image Generation
arXiv:2608.14172v1 Announce Type: cross Abstract: Text-to-image diffusion models have two major drawbacks that severely limit their practical utility: (1) stand
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper 1mo ago
DUET: A Diversity-Quality Duet of Distillation Experts for Two-Step Video Generation
arXiv:2608.09637v1 Announce Type: cross Abstract: Diffusion models have enabled high-quality video generation in recent years, but the high cost of iterative sa
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 1mo ago
FlowForm: Synergizing Fluid Physics with Topological Consistency for Satellite Flood Synthesis
arXiv:2608.03822v1 Announce Type: cross Abstract: Developing robust flood assessment models requires high-quality paired satellite imagery, yet such data remain
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 1mo ago
Element-Aware Group Learning for E-Commerce Image Generation
arXiv:2608.00584v1 Announce Type: cross Abstract: Recent advances in image generation and editing have made prompt quality a key bottleneck for e-commerce creat
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
A Systematic Survey on Image Description Techniques for STEM Domains
arXiv:2607.21611v1 Announce Type: cross Abstract: The proliferation of visual data in Science, Technology, Engineering, and Mathematics (STEM) fields presents a
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
Kontinuous Kontext: Continuous Strength Control for Instruction-based Image Editing
arXiv:2510.08532v2 Announce Type: replace-cross Abstract: Instruction-based image editing offers a powerful and intuitive way to manipulate images through natur
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
EmoStyle: Affective Conditioning of Style-Specialist Experts for Emotional Image Generation
arXiv:2607.10165v1 Announce Type: cross Abstract: Emotion-aware artistic image generation requires an image to match the input prompt, follow the specified arti
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
NI-Tex: Non-isometric Image-based Garment Texture Generation
arXiv:2511.18765v3 Announce Type: replace-cross Abstract: Existing industrial 3D garment meshes already cover most real-world clothing geometries, yet their tex
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
Histogram-constrained Image Generation
arXiv:2606.31683v1 Announce Type: cross Abstract: Diffusion models have emerged as a dominant paradigm in generative modeling, enabling high-fidelity sampling f
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
Learning to Adaptively Allocate Gaussians for Arbitrary-Scale Image Super-Resolution
arXiv:2606.29400v1 Announce Type: cross Abstract: In computer graphics, visual content is continuously warped, zoomed and resampled. This occurs when engines up
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
Resonant Brane Splatting for Arbitrary-Scale Super-Resolution
arXiv:2606.29453v1 Announce Type: cross Abstract: Arbitrary-Scale Super-Resolution (ASR) reconstructs images at continuous magnification factors. Recent methods
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
Home3D 1.0: A High-Fidelity Image-to-3D Asset Generation System for Interior Design
arXiv:2606.27923v1 Announce Type: cross Abstract: We present Home3D 1.0, a modular image-to-3D generation system that produces high-quality 3D assets from a sin
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
BiDeMem: Bidirectional Degradation Memory for Explainable Image Restoration
arXiv:2606.28112v1 Announce Type: cross Abstract: Degradation-aware prompts, conditions, and latent priors are increasingly used in image restoration, yet they
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 2mo ago
Deepfake Media Generation and Detection in the Generative AI Era: A Survey and Outlook
arXiv:2411.19537v2 Announce Type: replace-cross Abstract: We survey deepfake generation and detection techniques, covering all deepfake media types: image, vide
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 3mo ago
Scaling Multi-Reference Image Generation with Dynamic Reward Optimization
arXiv:2606.26947v1 Announce Type: cross Abstract: While personalized image generation has achieved remarkable progress, multi-reference image generation (MRIG)
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 3mo ago
Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks
arXiv:2606.27147v1 Announce Type: cross Abstract: Unlike diffusion-based models that operate in continuous latent spaces, autoregressive unified multimodal mode
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 3mo ago
BELDE: Building a Large-scale Earth-observation Land-cover Dataset for Europe
arXiv:2606.20909v1 Announce Type: cross Abstract: Earth observation imagery plays a critical role in environmental monitoring, urban planning, disaster assessme
ArXiv cs.AI 🎨 Image & Video AI 📄 Paper ⚡ AI Lesson 3mo ago
Text-to-Image Generative AI for Modeling and Simulation: Methods, Opportunities, and Applications
arXiv:2606.20991v1 Announce Type: cross Abstract: Text-to-image generation is a form of generative artificial intelligence (GenAI) that converts textual descrip