
research note
Archon — A Unified Multimodal Model for Holistic Digital Human Generation
This paper introduces Archon, a fully pretrained unified multimodal model designed specifically for holistic digital human generation

research note
This paper introduces Archon, a fully pretrained unified multimodal model designed specifically for holistic digital human generation

research note
This paper addresses the problem of embedding high-capacity watermarks into 3D Gaussian Splatting (3DGS) assets to enable robust ownership and provenance verification in large-scale 3D content pipe…

research note
This paper addresses privacy, availability, and authenticity challenges in Austria's national electronic identity (eID) system pseudonyms, called bPks, which rely on a fully centralized architecture

research note
This paper addresses critical limitations of centralized biometric identity systems, such as single points of failure, opaque verification, and irreversible biometric compromise

research note
City-Mesh3R addresses the challenge of reconstructing high-fidelity, watertight 3D meshes at city scale from large, unordered collections of multi-view images

research note
This paper addresses the challenge of incorporating informal, real-world Korean language expressions found in web documents into Korean e-learning systems targeted at high-level learners

research note
CommunityFact introduces a dynamic, multilingual, and multi-domain benchmark for misinformation detection that better reflects real-world, ongoing misinformation verification challenges on social m…

research note
This paper addresses the challenge of generating plausible future mathematical claims that are both scientifically motivated and logically valid

research note
This paper addresses the challenging problem of evaluating interactive front-end web code generated by large language models (LLMs)

research note
This paper addresses the underexplored but important problem of how the temporal organization of training data influences Large Language Model (LLM) training dynamics and final performance

research note
This paper introduces DiffSpot, a benchmark designed to evaluate vision-language models (VLMs) on the fine-grained visual perception task of spotting subtle differences in rendered web interfaces

research note
This paper addresses the foundational challenge of establishing trust and credible reputation systems for autonomous language model (LM) agents as they proliferate and interact in increasingly cons…