Exploring the Potential of Contrastive Language-Image Pre-training for Multi-Source Remote Sensing Data
Beyond Execution: Auditing Experimental Fidelity in LLM-Driven Scientific Research
LLM-based Agents for Forecasting and Prediction: Methods, Training, Evaluation, and Applications
ARAC: Benchmarking Auto-Research's Alignment and Completeness on End-to-End Researchs
Reasoning Denoiser: Denoising Reasoning Traces for Hallucination Detection in Large Reasoning Models
HDRAgent: An Agentic Framework for Multi-Exposure HDR Imaging
LITMUS: Benchmarking Behavioral Jailbreaks of LLM Agents in Real OS Environments
AHPA: Adaptive Hierarchical Prior Alignment for Diffusion Transformers
From Thinker to Society: Security in Hierarchical Autonomy Evolution of AI Agents
Low-Light Video Enhancement with An Effective Spatial-Temporal Decomposition Paradigm
MiroThinker-1.7 & H1: Towards Heavy-Duty Research Agents via Verification
Multi-Agent Tool-Integrated Policy Optimization with Process Reward
Enhancing Diffusion-based Restoration Models via Difficulty-Adaptive Reinforcement Learning with IQA Reward
Boosting Fidelity for Pre-Trained-Diffusion-Based Low-Light Image Enhancement via Condition Refinement
MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling
Jailbreaking Commercial Black-Box LLMs with Explicitly Harmful Prompts
Generative Distribution Distillation
Generative Visual Commonsense Answering and Explaining with Generative Scene Graph Constructing
Iter-AHMCL: Alleviate Hallucination for Large Language Model via Iterative Model-level Contrastive Learning
Self-supervised Learning for Enhancing Geometrical Modeling in 3D-Aware Generative Adversarial Network
Clarity ChatGPT: An Interactive and Adaptive Processing System for Image Restoration and Enhancement
General Adversarial Defense Against Black-box Attacks via Pixel Level and Feature Level Distribution Alignments