AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,584 results
22 Apr 2026

MOSA: Motion-Guided Semantic Alignment for Dynamic Scene Graph Generation

SafetyDGX agent

arXiv:2604.19631v1 Announce Type: new Abstract: Dynamic Scene Graph Generation (DSGG) aims to structurally model objects and their dynamic interactions in video sequences for high-level semantic under

Multiclass Local Calibration with the Jensen-Shannon Distance

SafetyDGX agent

arXiv:2510.26566v2 Announce Type: replace-cross Abstract: Developing trustworthy Machine Learning (ML) models requires their predicted probabilities to be well-calibrated, meaning they should reflect

NemeSys: Toward Online Underwater Exploration with Remote Operator-in-the-loop Adaptive Autonomy

Model ReleasesDGX agent

arXiv:2507.11889v2 Announce Type: replace Abstract: Adaptive mission control and dynamic parameter reconfiguration are essential for autonomous underwater vehicles (AUVs) operating in GPS-denied, comm

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Nexusformer: Nonlinear Attention Expansion for Stable and Inheritable Transformer Scaling

ResearchDGX agent

arXiv:2604.19147v1 Announce Type: cross Abstract: Scaling Transformers typically necessitates training larger models from scratch, as standard architectures struggle to expand without discarding learn

On Temperature-Constrained Non-Deterministic Machine Translation: Potential and Evaluation

ApplicationsDGX agent

arXiv:2601.13729v2 Announce Type: replace Abstract: In recent years, the non-deterministic properties of language models have garnered considerable attention and have shown a significant influence on

On the Conditioning Consistency Gap in Conditional Neural Processes

ResearchDGX agent

arXiv:2604.19312v1 Announce Type: new Abstract: Neural processes are meta-learning models that map context sets to predictive distributions. While inspired by stochastic processes, NPs do not generall

Pixels or Positions? Benchmarking Modalities in Group Activity Recognition

Model ReleasesDGX agent

arXiv:2511.12606v3 Announce Type: replace Abstract: Group Activity Recognition (GAR) is well studied on the video modality for surveillance and indoor team sports (e.g., volleyball, basketball). Yet,

PolarQuant: Optimal Gaussian Weight Quantization via Hadamard Rotation for LLM Compression

Local AiDGX agent

arXiv:2603.29078v2 Announce Type: replace Abstract: We present PolarQuant, a post-training weight quantization method for large language models (LLMs) that exploits the distributional structure of neu

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infra…

SafetyDGX agent

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infrastructure & AI software that underpins our Robotaxi & future

RAFT-MSF++: Temporal Geometry-Motion Feature Fusion for Self-Supervised Monocular Scene Flow

Model ReleasesDGX agent

arXiv:2604.19349v1 Announce Type: new Abstract: Monocular scene flow estimation aims to recover dense 3D motion from image sequences, yet most existing methods are limited to two-frame inputs, restric

RESFL: An Uncertainty-Aware Framework for Responsible Federated Learning by Balancing Privacy, Fairness and Utility

SafetyDGX agent

arXiv:2503.16251v2 Announce Type: replace-cross Abstract: Federated Learning (FL) has gained prominence in machine learning applications across critical domains by enabling collaborative model trainin

SAMoRA: Semantic-Aware Mixture of LoRA Experts for Task-Adaptive Learning

Model ReleasesDGX agent

arXiv:2604.19048v1 Announce Type: cross Abstract: The combination of Mixture-of-Experts (MoE) and Low-Rank Adaptation (LoRA) has shown significant potential for enhancing the multi-task learning capab

See2Refine: Vision-Language Feedback Improves LLM-Based eHMI Action Designers

ResearchDGX agent

arXiv:2602.02063v2 Announce Type: replace-cross Abstract: Automated vehicles lack natural communication channels with other road users, making external Human-Machine Interfaces (eHMIs) essential for c

Semantic Needles in Document Haystacks: Sensitivity Testing of LLM-as-a-Judge Similarity Scoring

SafetyDGX agent

arXiv:2604.18835v1 Announce Type: cross Abstract: We propose a scalable, multifactorial experimental framework that systematically probes LLM sensitivity to subtle semantic changes in pairwise documen

Separating Geometry from Probability in the Analysis of Generalization

ResearchDGX agent

arXiv:2604.19560v1 Announce Type: new Abstract: The goal of machine learning is to find models that minimize prediction error on data that has not yet been seen. Its operational paradigm assumes acces

SketchFaceGS: Real-Time Sketch-Driven Face Editing and Generation with Gaussian Splatting

ResearchDGX agent

arXiv:2604.19202v1 Announce Type: cross Abstract: 3D Gaussian representations have emerged as a powerful paradigm for digital head modeling, achieving photorealistic quality with real-time rendering.

Small and midsize businesses jumpstart their AI transformations with Gemini Enterprise

Model ReleasesDGX agent

Small businesses are the backbone of the global economy. With 400 million SMBs worldwide and 36 million in the U.S. alone, they provide 50% of global employment. Now, with Google Cloud AI, they’re sca

SmokeGS-R: Physics-Guided Pseudo-Clean 3DGS for Real-World Multi-View Smoke Restoration

Model ReleasesDGX agent

arXiv:2604.05301v2 Announce Type: replace Abstract: Real-world smoke simultaneously attenuates scene radiance, adds airlight, and destabilizes multi-view appearance consistency, making robust 3D recon

SPRITE: From Static Mockups to Engine-Ready Game UI

Model ReleasesDGX agent

arXiv:2604.18591v1 Announce Type: cross Abstract: Game UI implementation requires translating stylized mockups into interactive engine entities. However, current 'Screenshot-to-Code' tools often strug

Stable-RAG: Mitigating Retrieval-Permutation-Induced Hallucinations in Retrieval-Augmented Generation

SafetyDGX agent

arXiv:2601.02993v4 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) has become a key paradigm for reducing factual hallucinations in Large Language Models (LLMs), yet little is kn

Thank you for flagging this, Jeff. This was a mistake: we are not deprecating text-embedding-3-small. We’re looking into where this came fro…

ToolsDGX agent

Thank you for flagging this, Jeff. This was a mistake: we are not deprecating text-embedding-3-small. We’re looking into where this came from now, and we’ll also email users to clarify. Sorry for the

The decisive layer in AI is still unclaimed: theCUBE’s Google Cloud Next day one keynote analysis

Model ReleasesDGX agent

The fight for the agent control plane is underway — and it might determine who controls enterprise artificial intelligence for the next decade. Google LLC came into Google Cloud Next 2026 with a clear

The future of data lakehouse: Open and interoperable for the agentic era

Model ReleasesDGX agent

Traditional lakehouses were engineered for the era of reporting, not the high-velocity, multimodal demands of AI agents. To bridge this gap, architecture must evolve into an AI-native foundation — one

Toward Clinically Acceptable Chest X-ray Report Generation: A Qualitative Retrospective Pilot Study of CXRMate-2

SafetyDGX agent

arXiv:2604.18967v1 Announce Type: new Abstract: Chest X-ray (CXR) radiology report generation (RRG) models have shown rapid progress, yet their clinical utility remains uncertain due to limited evalua

Towards Optimal Agentic Architectures for Offensive Security Tasks

Model ReleasesDGX agent

arXiv:2604.18718v1 Announce Type: cross Abstract: Agentic security systems increasingly audit live targets with tool-using LLMs, but prior systems fix a single coordination topology, leaving unclear w

Unposed-to-3D: Learning Simulation-Ready Vehicles from Real-World Images

AgentsDGX agent

arXiv:2604.19257v1 Announce Type: new Abstract: Creating realistic and simulation-ready 3D assets is crucial for autonomous driving research and virtual environment construction. However, existing 3D

Unsupervised Confidence Calibration for Reasoning LLMs from a Single Generation

ResearchDGX agent

arXiv:2604.19444v1 Announce Type: new Abstract: Reasoning language models can solve increasingly complex tasks, but struggle to produce the calibrated confidence estimates necessary for reliable deplo

URoPE: Universal Relative Position Embedding across Geometric Spaces

Model ReleasesDGX agent

arXiv:2604.18747v1 Announce Type: new Abstract: Relative position embedding has become a standard mechanism for encoding positional information in Transformers. However, existing formulations are typi

VISTA: Verification In Sequential Turn-based Assessment

ResearchDGX agent

arXiv:2510.27052v5 Announce Type: replace Abstract: Hallucination--defined here as generating statements unsupported or contradicted by available evidence or conversational context--remains a major ob

Whispers in the Machine: Confidentiality in Agentic Systems

AgentsDGX agent

arXiv:2402.06922v5 Announce Type: replace-cross Abstract: Large language model (LLM)-based agents combine LLMs with external tools to automate tasks such as scheduling meetings, managing documents, or

Why Does Self-Distillation (Sometimes) Degrade the Reasoning Capability of LLMs?

Model ReleasesDGX agent

arXiv:2603.24472v2 Announce Type: replace Abstract: Self-distillation has emerged as an effective post-training paradigm for LLMs, often improving performance while shortening reasoning traces. Howeve

With Gemini Enterprise Agent Platform, Google brings agentic development and control under one roof

Model ReleasesDGX agent

Google Cloud is taking a massive leap toward building the autonomous enterprise with the launch of the Gemini Enterprise Agent Platform, an evolution of the existing Vertex AI platform that becomes it

21 Apr 2026

A Discordance-Aware Multimodal Framework with Multi-Agent Clinical Reasoning

AgentsDGX agent

arXiv:2604.16333v1 Announce Type: new Abstract: Knee osteoarthritis frequently exhibits discordance between structural damage observed in imaging and patient-reported symptoms such as pain. This misma

AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficient Kernel Generation

AgentsDGX agent

arXiv:2604.16625v1 Announce Type: new Abstract: Recent large language model (LLM) agents have shown promise in using execution feedback for test-time adaptation. However, robust self-improvement remai

AIM 2025 Rip Current Segmentation (RipSeg) Challenge Report

Model ReleasesDGX agent

arXiv:2508.13401v3 Announce Type: replace Abstract: This report presents an overview of the AIM 2025 RipSeg Challenge, a competition designed to advance techniques for automatic rip current segmentati

Alexandria: A Multi-Domain Dialectal Arabic Machine Translation Dataset for Culturally Inclusive and Linguistically Diverse LLMs

Model ReleasesDGX agent

arXiv:2601.13099v2 Announce Type: replace Abstract: Arabic is a highly diglossic language where most daily communication occurs in regional dialects rather than Modern Standard Arabic (MSA). Despite t

Annotation Entropy Predicts Per-Example Learning Dynamics in LoRA Fine-Tuning

ResearchDGX agent

arXiv:2604.16332v1 Announce Type: cross Abstract: We find that LoRA fine-tuning exhibits un-learning on contested examples: items with high annotator disagreement show increasing loss during training,

Arch: An AI-Native Hardware Description Language for Register-Transfer Clocked Hardware Design

SafetyDGX agent

arXiv:2604.05983v2 Announce Type: replace-cross Abstract: We present Arch (AI-native Register-transfer Clocked Hardware), a hardware description language for micro-architecture specification and AI-as

Are they lovers or friends? Evaluating LLMs' Social Reasoning in English and Korean Dialogues

ApplicationsDGX agent

arXiv:2510.19028v3 Announce Type: replace Abstract: As LLMs are increasingly deployed in real-world interactions, their social reasoning in interpersonal communication becomes critical. To explore the

Attention-space Contrastive Guidance for Efficient Hallucination Mitigation in LVLMs

SafetyDGX agent

arXiv:2601.13707v2 Announce Type: replace Abstract: Hallucinations in large vision--language models (LVLMs) often arise when language priors dominate over visual evidence, leading to object misidentif

Back into Plato's Cave: Examining Cross-modal Representational Convergence at Scale

SafetyDGX agent

arXiv:2604.18572v1 Announce Type: new Abstract: The Platonic Representation Hypothesis suggests that neural networks trained on different modalities (e.g., text and images) align and eventually conver

BASIL: Bayesian Assessment of Sycophancy in LLMs

ApplicationsDGX agent

arXiv:2508.16846v5 Announce Type: replace-cross Abstract: Sycophancy (overly agreeable or flattering behavior) poses a fundamental challenge for human-AI collaboration, particularly in high-stakes dec

Beyond Feature Fusion: Contextual Bayesian PEFT for Multimodal Uncertainty Estimation

Model ReleasesDGX agent

arXiv:2604.16615v1 Announce Type: new Abstract: We introduce CoCo-LoRA, a multimodal, uncertainty-aware parameter-efficient fine-tuning method for text prediction tasks accompanied by audio context. E

Beyond Fine-Tuning: In-Context Learning and Chain-of-Thought for Reasoned Distractor Generation

ResearchDGX agent

arXiv:2604.17574v1 Announce Type: new Abstract: Distractor generation (DG) remains a labor-intensive task that still significantly depends on domain experts. The task focuses on generating plausible y

Beyond Overlap Metrics: Rewarding Reasoning and Preferences for Faithful Multi-Role Dialogue Summarization

SafetyDGX agent

arXiv:2604.17188v1 Announce Type: new Abstract: Multi-role dialogue summarization requires modeling complex interactions among multiple speakers while preserving role-specific information and factual

Beyond Pattern Matching: Seven Cross-Domain Techniques for Prompt Injection Detection

Model ReleasesDGX agent

arXiv:2604.18248v1 Announce Type: cross Abstract: Current open-source prompt-injection detectors converge on two architectural choices: regular-expression pattern matching and fine-tuned transformer c

Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining

ResearchDGX agent

arXiv:2511.21613v2 Announce Type: replace Abstract: Incorporating metadata in Large Language Models (LLMs) pretraining has recently emerged as a promising approach to accelerate training. However prio

BhashaSutra: A Task-Centric Unified Survey of Indian NLP Datasets, Corpora, and Resources

ResearchDGX agent

arXiv:2604.18423v1 Announce Type: new Abstract: India's linguistic landscape, spanning 22 scheduled languages and hundreds of marginalized dialects, has driven rapid growth in NLP datasets, benchmarks

Brain-Inspired Capture: Evidence-Driven Neuromimetic Perceptual Simulation for Visual Decoding

Model ReleasesDGX agent

arXiv:2604.17927v1 Announce Type: new Abstract: Visual decoding of neurophysiological signals is a critical challenge for brain-computer interfaces (BCIs) and computational neuroscience. However, curr

CAM3DNet: Comprehensively mining the multi-scale features for 3D Object Detection with Multi-View Cameras

Model ReleasesDGX agent

arXiv:2604.17024v1 Announce Type: new Abstract: Query-based 3D object detection methods using multi-view images often struggle to efficiently leverage dynamic multi-scale information, e.g., the relati

Camo-M3FD: A New Benchmark Dataset for Cross-Spectral Camouflaged Pedestrian Detection

Model ReleasesDGX agent

arXiv:2604.16582v1 Announce Type: new Abstract: Pedestrian detection is fundamental to autonomous driving, robotics, and surveillance. Despite progress in deep learning, reliable identification remain

CDSA-Net:Collaborative Decoupling of Vascular Structure and Background for High-Fidelity Coronary Digital Subtraction Angiography

Model ReleasesDGX agent

arXiv:2604.17208v1 Announce Type: new Abstract: Digital subtraction angiography (DSA) in coronary imaging is fundamentally challenged by physiological motion, forcing reliance on raw angiograms clutte

CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark

Model ReleasesDGX agent

arXiv:2604.16372v1 Announce Type: new Abstract: Multimodal sarcasm detection has recently garnered significant attention. However, existing benchmarks suffer from coarse-grained annotations and limite

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt ha…

Model ReleasesDGX agent

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt hacks like 'think step by step' don't count here, I mean the e

Co-generation of Layout and Shape from Text via Autoregressive 3D Diffusion

ResearchDGX agent

arXiv:2604.16552v1 Announce Type: new Abstract: Recent text-to-scene generation approaches largely reduced the manual efforts required to create 3D scenes. However, their focus is either to generate a

CoDial: Interpretable Task-Oriented Dialogue Systems Through Dialogue Flow Alignment

Model ReleasesDGX agent

arXiv:2506.02264v3 Announce Type: replace Abstract: Building Task-Oriented Dialogue (TOD) systems that generalize across different tasks remains a challenging problem. Data-driven approaches often str

Cognitive Policy-Driven LLM for Diagnosis and Intervention of Cognitive Distortions in Emotional Support Conversation

SafetyDGX agent

arXiv:2604.17178v1 Announce Type: new Abstract: Emotional Support Conversation (ESC) plays a critical role in mental health assistance by providing accessible psychological support in real-world appli

Context Matters: Peer-Aware Student Behavioral Engagement Measurement via VLM Action Parsing and LLM Sequence Classification

ResearchDGX agent

arXiv:2601.06394v3 Announce Type: replace Abstract: Understanding student behavior in the classroom is essential to improve both pedagogical quality and student engagement. Existing methods for predic

Data Compressibility Quantifies LLM Memorization

ResearchDGX agent

arXiv:2507.06056v4 Announce Type: replace Abstract: Large Language Models (LLMs) are known to memorize portions of their training data, sometimes even reproduce content verbatim when prompted appropri

Decoding RWA Tokenized U.S. Treasuries: Functional Dissection and Address Role Inference

ApplicationsDGX agent

arXiv:2507.14808v3 Announce Type: replace-cross Abstract: Tokenized U.S. Treasuries have emerged as a prominent subclass of real-world assets (RWAs), offering cryptographically secured, yield-bearing

← Previous
1…576577578579580…1060
Next →