AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
3 Jul 2026

CoRe: Combined Rewards with Vision-Language Model Feedback for Preference-Aligned Reinforcement Learning

SafetyDGX agent

arXiv:2607.01721v1 Announce Type: new Abstract: Reward design remains a central challenge in reinforcement learning (RL). Hand-crafted rewards are often difficult to specify and may lead to suboptimal

Epistemic Goggles: A Pretrained Module that Induces an Epistemic Frame via Gradient Editing

SafetyDGX agent

arXiv:2607.01690v1 Announce Type: new Abstract: Finetuning a language model on documents that are explicitly annotated as fictional results in a model that still actually believes the documents' core

Exploring Large Language Models for Access Control Policy Synthesis and Summarization

SafetyDGX agent

arXiv:2510.20692v2 Announce Type: replace-cross Abstract: Cloud computing is ubiquitous, with a growing number of services being hosted on the cloud every day. Typical cloud compute systems allow admi

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

G-RRM: Guiding Symbolic Solvers with Recurrent Reasoning Models

TutorialsDGX agent

arXiv:2607.02491v1 Announce Type: new Abstract: In this work, we focus on SE-RRMs, a symbol-equivariant instantiation of RRMs that exhibits improved extrapolation to larger problem sizes. We propose a

Generative AI and Federated Learning for Intrusion Detection Systems: A Survey

Model ReleasesDGX agent

arXiv:2607.01305v1 Announce Type: cross Abstract: Intrusion Detection Systems (IDSs) are essential for monitoring network traffic and identifying malicious activities in modern cyber-physical, Interne

HarnessX: A Composable, Adaptive, and Evolvable Agent Harness Foundry

Model ReleasesDGX agent

arXiv:2606.14249v2 Announce Type: replace Abstract: AI agent performance depends critically on the runtime harness, comprising the prompts, tools, memory, and control flow that mediate how a model obs

OpenAI offers feds a stake, Anthropic gets out of AI model jail and Meta wants to be a neocloud

SafetyDGX agent

OpenAI reportedly has floated giving the U.S. government a 5% stake in the company, perhaps the start of a series of such stakes in other AI companies as well. This no doubt has traditional anti-indus

Robust for the Wrong Reasons: The Representational Geometry of LLM Robustness to Science Skepticism

Model ReleasesDGX agent

arXiv:2607.01951v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly consulted on contested scientific questions, raising the concern that they will sycophantically retreat

Safety Targeted Embedding Exploit via Refinement

Model ReleasesDGX agent

arXiv:2607.01859v1 Announce Type: new Abstract: Safety training for large language models (LLMs) is conducted predominantly in English, leaving uncertain how well safety mechanisms generalize to low-r

Towards Robustness against Typographic Attack with Training-free Concept Localization

Model ReleasesDGX agent

arXiv:2607.02494v1 Announce Type: cross Abstract: Models trained via Contrastive Language-Image Pretraining (CLIP) serve as the foundational vision encoders for most modern Large Vision Language Model

WARP: Weight-Space Analysis for Recovering Training Data Portfolios

Model ReleasesDGX agent

arXiv:2607.01686v1 Announce Type: new Abstract: Foundation models are routinely released to the public, yet the data recipes used to train them -- such as domain mixture weights that determine how dif

World Wide Models: Literary Tools for Cultural AI

ResearchDGX agent

arXiv:2607.02369v1 Announce Type: cross Abstract: LLMs stage a new form of cultural encounter that is massive, automated, and monolingual. Literary disciplines have always negotiated cultural struggle

2 Jul 2026

A Geometric Perspective on Composable Emotion Steering in Text-to-Speech Models

ResearchDGX agent

arXiv:2607.00946v1 Announce Type: cross Abstract: While prior work has explored emotion control in hybrid text-to-speech systems, the geometric properties of these modules, and their implications for

Adversarial Pragmatics for AI Safety Evaluation: A Benchmark for Instruction Conflict, Embedded Commands, and Policy Ambiguity

Model ReleasesDGX agent

arXiv:2607.01153v1 Announce Type: cross Abstract: Safety evaluations for language models increasingly depend on judgments about ambiguous natural-language behaviour: whether a model has followed an in

BaseRT: Best-in-Class LLM Inference on Apple Silicon via Native Metal

Model ReleasesDGX agent

arXiv:2607.00501v1 Announce Type: cross Abstract: We present BaseRT, a native Metal inference runtime for large language models (LLMs) on Apple Silicon, and report the highest inference throughput on

BrainFIBRE: A Foundation Model via Information Decomposition for Brain Microstructure

SafetyDGX agent

arXiv:2607.00573v1 Announce Type: new Abstract: Diffusion MRI probes brain microstructure with particular sensitivity to early cerebrovascular and neurodegenerative changes. Neurite Orientation Disper

Bridgewater just published numbers that should make every frontier lab nervous. The world's largest hedge fund tested Gemini, Claude, and GP…

Model ReleasesDGX agent

Bridgewater just published numbers that should make every frontier lab nervous. The world's largest hedge fund tested Gemini, Claude, and GPT on six document filtering tasks its investors do every day

FastBridge: Closing the Model-Based Realization Gap in Safety Filters on 3D Gaussian Splatting for Fast Quadrotor Flight

SafetyDGX agent

arXiv:2607.01200v1 Announce Type: new Abstract: Fast quadrotor flight requires safe obstacle avoidance under tight onboard compute limits. While 3D Gaussian Splatting (3DGS) provides a continuous, geo

FurnitureVLA: Learning Long-Horizon Bimanual Furniture Assembly with Vision-Language-Action Model

ApplicationsDGX agent

arXiv:2607.01212v1 Announce Type: cross Abstract: Current work on robot furniture assembly mostly focuses on toy-scale settings or single-arm manipulation. We introduce FurnitureVLA, the first systema

FusionFactory: Fusing LLM Capabilities with Multi-LLM Log Data

Model ReleasesDGX agent

arXiv:2507.10540v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has created a diverse landscape of models, each excelling at different tasks. This diversity d

Is One Layer Enough? Training A Single Transformer Layer Can Match Full-Parameter RL Training

Model ReleasesDGX agent

arXiv:2607.01232v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central component of post-training large language models (LLMs), yet little is understood about how RL adapta

Neural Surface and Reflectance Modelling from 3D Radar Data

AgentsDGX agent

arXiv:2603.25623v2 Announce Type: replace Abstract: Robust scene representation is essential for autonomous systems to safely operate in challenging low-visibility environments. In these conditions, r

OpenReward: Learning to Reward Long-form Agentic Tasks via Reinforcement Learning

SafetyDGX agent

arXiv:2510.24636v3 Announce Type: replace Abstract: Reward models (RMs) have become essential for aligning large language models (LLMs), serving as scalable proxies for human evaluation in both traini

PAPA: Online Personalized Active Preference Alignment

SafetyDGX agent

arXiv:2607.00486v1 Announce Type: cross Abstract: Diffusion models are highly effective at modeling complex data distributions, including images and text. However, in applications like personalized re

Personalized Object Identification and Localization via In-Context Inference with Vision-Language Models

Local AiDGX agent

arXiv:2607.00357v1 Announce Type: new Abstract: Personalized object localization (POL) localizes an object instance in a query image based on a few reference images with bounding-box annotations and a

Spatio-Temporal Gaussian Process for Building Terrain-Incorporating Wind Power Curves

SafetyDGX agent

arXiv:2607.00051v1 Announce Type: cross Abstract: Accurate modeling of wind turbine power curves is crucial for optimal wind farm operation. Nearly all existing power curve models focus on temporal va

SpiralFovea: Input-Adaptive Foveated Tokenization as a Third Lever of Resource-Adaptive Inference

Model ReleasesDGX agent

arXiv:2607.00780v1 Announce Type: new Abstract: Most adaptive-inference techniques for foundation models change what the model does - early exit, MoE routing, KV-cache compression, dynamic attention s

StateFlow: Dual-State Recurrent Modeling for Long-Horizon Time Series Forecasting

Local AiDGX agent

arXiv:2607.00197v1 Announce Type: new Abstract: Long-horizon multivariate time series forecasting (LTSF) remains challenging due to non-stationarity, regime shifts, and error accumulation. The Variabi

The MMM Data Model -- A Normative Specification for Knowledge Interoperability in a Decentralisable Knowledge Commons

ApplicationsDGX agent

arXiv:2607.00032v1 Announce Type: new Abstract: Many information systems are built around documents: self-contained units optimised for print production and linear reading. While effective for large-s

1 Jul 2026

CharDiff-LP: A Diffusion Model with Character-Level Guidance for License Plate Image Restoration

ResearchDGX agent

arXiv:2510.17330v3 Announce Type: replace-cross Abstract: License plate image restoration is important not only as a preprocessing step for license plate recognition but also for enhancing evidential

DSIP: A Dynamic Coordination Planner for Signal-Free Intersections using Diffusion-Model-Based Multi-Agent Motion Planning

AgentsDGX agent

arXiv:2606.30694v1 Announce Type: cross Abstract: Traffic signal control at urban intersections inherently introduces stop-and-go behavior, resulting in increased delays and reduced traffic efficiency

Localized Conformal Prediction for Image Classification with Vision-Language Models

ResearchDGX agent

arXiv:2606.31577v1 Announce Type: new Abstract: Conformal predictions have attracted significant attention in the field of uncertainty quantification, mainly because of their strong marginal coverage

Minimizing Quantized Semantic Age of Information (QSAoI) in Foundation Model-Based Semantic Communications

ResearchDGX agent

arXiv:2606.31303v1 Announce Type: cross Abstract: The emerging techniques of semantic communications and edge computing in 6G networks necessitate a paradigm shift toward co-designed semantic-aware an

Optimal Self-Consistency for Efficient Reasoning with Large Language Models

ResearchDGX agent

arXiv:2511.12309v2 Announce Type: replace-cross Abstract: Self-consistency (SC) is a widely used test-time inference technique for improving performance in chain-of-thought reasoning. It consists of g

RCL-Mamba: A Dual-domain State Space Model for Measurement-oriented Image Restoration in Rotational Sparse-View Scanning Computed Laminography

ApplicationsDGX agent

arXiv:2606.31353v1 Announce Type: new Abstract: Rotational Scanning Computed Laminography (RCL) is widely utilized for the Non-Destructive Testing(NDT) of large planar components. However, to facilita

Structure-Regularized Interpretable TCR-Epitope Prediction

Model ReleasesDGX agent

arXiv:2606.30902v1 Announce Type: cross Abstract: T cell receptor (TCR)-epitope binding prediction is essential for understanding adaptive immunity and developing immunotherapies. Existing sequence- a

The Consistency Dilemma in LLMs: Generator-Evaluator Agreement and Vulnerability to Mistakes

AgentsDGX agent

arXiv:2606.30653v1 Announce Type: cross Abstract: Large language models are increasingly deployed in agentic pipelines that depend on the model evaluating its own outputs without external verification

🔬 The Coolest Diffusion Research Isn't in LLMs — Evan Feinberg & Sergey Edunov, Genesis Molecular AI

Model ReleasesDGX agent

This episode discusses cutting-edge diffusion model research applications beyond large language models, featuring insights from Evan Feinberg and Sergey Edunov of Genesis Molecular AI on how diffusion

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action

Model ReleasesDGX agent

arXiv:2606.31916v1 Announce Type: new Abstract: Theory of Mind (ToM) benchmarks for Large Language Models (LLMs) typically rely on passive question-answering formats, but the deployment of LLMs in inc

Think in English, Answer in Korean: Efficient Adaptation of Multilingual Tool-Using Agents

Model ReleasesDGX agent

arXiv:2606.31648v1 Announce Type: new Abstract: We present LuckyStar 111B, a 111B-parameter hybrid reasoning model developed through a collaboration between Cohere and LG CNS for Korean-English enterp

Towards a foundational model for recognising diastematic Gregorian notation

ResearchDGX agent

arXiv:2606.31454v1 Announce Type: new Abstract: Optical recognition of Gregorian notation has recently been attempted with end-to-end methods, with four datasets introduced. However, each of these dat

TreeAgent: A Generalizable Multi-Agent Framework for Automated Bias Labeling in Forestry via Compiled Expert Rules and Vision-Language Models

SafetyDGX agent

arXiv:2606.31976v1 Announce Type: new Abstract: Human-labeled data are widely used as reference annotations in ML, despite known variability across annotators in many expert-driven domains. In additio

Vision-Language Procedural Reasoning for Context-Aware Reward Modeling of Robotic Endovascular Guidewire Navigation

SafetyDGX agent

arXiv:2606.30698v1 Announce Type: new Abstract: Robotic-assisted endovascular interventions demand accurate, stable, and context-aware guidewire navigation in complex and patient-specific vascular ana

When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs

ResearchDGX agent

arXiv:2606.30814v1 Announce Type: new Abstract: Calibration evaluates whether a model confidence aligns with its empirical accuracy. Existing studies often compare the calibration of different large l

When LLMs Read Tables Carelessly: Measuring and Reducing Data Referencing Errors

Model ReleasesDGX agent

arXiv:2606.32029v1 Announce Type: cross Abstract: While large language models (LLMs) perform well on table tasks, they still make data referencing errors (DREs), i.e., incorrectly citing or omitting t

When Sinks Help or Hurt: Unified Framework for Attention Sink in Large Vision-Language Models

ResearchDGX agent

arXiv:2604.03316v2 Announce Type: replace Abstract: Attention sinks are defined as tokens that attract disproportionate attention. While these have been studied in single modality transformers, their

When the Database Fails: Prompting LLM Dialogue Agents for Safe Recovery in Task-Oriented Dialogue

Model ReleasesDGX agent

arXiv:2606.31307v1 Announce Type: new Abstract: Large language models used in task-oriented dialogue often produce fluent but unsafe responses when backend database calls fail, return empty results, o

30 Jun 2026

AEGIR: Modeling Area Emitters for Indoor Inverse Rendering using Gaussian Splatting

ApplicationsDGX agent

arXiv:2606.28635v1 Announce Type: new Abstract: Inverse rendering requires separating illumination from surface materials, which is highly ambiguous due to their tight coupling in observed images. Whi

AnTenA: Actionable and Explainable Tensor Analysis System with Large Language Models

ResearchDGX agent

arXiv:2606.28708v1 Announce Type: new Abstract: Accurately explaining hidden patterns in multi-aspect data has typically been done by leveraging labels and/or accompanying auxiliary metadata. However,

Audio-Visual Continual Test-Time Adaptation without Forgetting

Model ReleasesDGX agent

arXiv:2602.18528v2 Announce Type: replace Abstract: Audio-visual continual test-time adaptation involves continually adapting a source audio-visual model at test-time, to unlabeled non-stationary doma

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning

TutorialsDGX agent

arXiv:2606.30217v1 Announce Type: new Abstract: Large multimodal models have achieved strong reasoning on complex visual tasks, but their inference efficiency is often restricted by long chains of tho

Beyond Scaling Law: A Data-Efficient Distillation Framework for Reasoning

Model ReleasesDGX agent

arXiv:2508.09883v2 Announce Type: replace-cross Abstract: Large language models (LLMs) demonstrate remarkable reasoning capabilities in tasks such as algorithmic coding and mathematical problem-solvin

Build agents even faster with Gemini Enterprise Agent Platform’s fully-managed, remote MCP server

Model ReleasesDGX agent

A couple of months ago, we announced that over 50 Google-managed MCP servers are available. Today, we’ll dive into how to use the Gemini Enterprise Agent Platform remote MCP server to securely connect

Closing the Activation-Cone Blind Spot: Response-Time Probing and Unified Defense

Model ReleasesDGX agent

arXiv:2606.29441v1 Announce Type: cross Abstract: Inference-time safety methods for large language models have proliferated, yet no systematic comparison exists. We evaluate five defense paradigms (no

DeVAR: Low-Dose CT Denoising via Visual Autoregressive Modeling

ResearchDGX agent

arXiv:2606.28453v1 Announce Type: cross Abstract: Computed tomography (CT) plays a crucial role in medical diagnosis, but minimizing radiation exposure while maintaining image quality remains a critic

Diff-Based Code Corruption using LLMs for Large-Scale Bugfix Benchmarking

Model ReleasesDGX agent

arXiv:2606.29088v1 Announce Type: cross Abstract: There are various benchmarks to evaluate bugfixing capabilities of Large Language Models. However, most widespread benchmarks do not fully reflect rea

Distribution Matching Variational AutoEncoder

SafetyDGX agent

arXiv:2512.07778v2 Announce Type: replace Abstract: Most visual generative models compress images into a latent space before applying diffusion or autoregressive modelling. Yet, existing approaches su

EMPATH: A Multilingual Auditor-Judge Benchmark for Safety Evaluation of Emotional-Support Chatbots

Model ReleasesDGX agent

arXiv:2606.30256v1 Announce Type: new Abstract: Safety benchmarks often buy scalability by fixing the prompt, the language, and the turn structure. For emotional-support chatbots, that bargain hides p

FinInvest-GTCN: Explainable Graph-Temporal-Causal Modeling for Risk-Aware Investment Decision Optimization

ResearchDGX agent

arXiv:2606.28933v1 Announce Type: new Abstract: Venture capital (VC) investment decisions face distinct challenges, such as multi-source heterogeneous data, non-stationary time series, and the demand

From Tool Connection to Execution Control: Benchmarking Security Invariants in MCP-Style Agent Runtimes

Model ReleasesDGX agent

arXiv:2606.29073v1 Announce Type: cross Abstract: Model Context Protocol (MCP)-style ecosystems give language-model applications a practical connection layer for tools, resources, prompts, and transpo

← Previous
1…253254255256257…1034
Next →