AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Research

Rethinking Uncertainty Evaluation in Large Language Models

DGX agent

arXiv:2607.19367v1 Announce Type: new Abstract: Calibration is the primary criterion for evaluating LLM confidence, but it is insufficient: it admits trivially incoherent estimators, depends on the ev

researcharxiv-cs-ai
23 Jul 2026
Model Releases
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RS-RIE-Bench: Benchmarking Reasoning-Guided Remote Sensing Image Editing

DGX agent

arXiv:2607.20197v1 Announce Type: new Abstract: Remote sensing image editing aims to modify remote sensing images according to natural language instructions while preserving geographic rules and senso

model-releasesarxiv-cs-cv
23 Jul 2026
Local Ai

Self-supervision drives representational convergence in medical foundation models more than clinical supervision

DGX agent

arXiv:2607.20274v1 Announce Type: cross Abstract: Medical image encoders from different groups are increasingly treated as interchangeable, on the assumption that scale and clinical supervision concen

local-aiarxiv-cs-ai
23 Jul 2026
Model Releases

When Shippers Become Algorithms: Candidate Exposure, Information Design, and the Concentration of LLM-Mediated Freight Markets

DGX agent

arXiv:2607.19967v1 Announce Type: cross Abstract: Shippers are beginning to delegate carrier selection to large language model (LLM) agents. We ask what such delegation does to a freight matching mark

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

DNA: Dual-stage Native Attribution for Generated Image Source Tracing

DGX agent

arXiv:2607.13685v1 Announce Type: new Abstract: The rapid evolution of image generation has produced numerous within-family variants, making source-model attribution of suspect images increasingly imp

model-releasesarxiv-cs-cv
16 Jul 2026
Safety

MASPRM: Multi-Agent System Process Reward Model

DGX agent

arXiv:2510.24803v3 Announce Type: replace-cross Abstract: Inference-time search over multi-agent systems (MAS) wastes compute when it cannot identify which agent's intermediate message advanced progre

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

Not All Needles Are Found: How Fact Distribution and Don't Make It Up Prompts Shape Retrieval, Reasoning, and Hallucination in Long-Context LLMs

DGX agent

arXiv:2601.02023v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) increasingly utilize massive context windows as working memory for autonomous tasks, their reliability fluctua

model-releasesarxiv-cs-ai
16 Jul 2026
Model Releases

Securing LLMs in the Wild: Privacy and Security Challenges at the Edge

DGX agent

arXiv:2607.13088v1 Announce Type: cross Abstract: Large Language Models (LLMs) are rapidly moving from research settings into the wild, deployed on enterprise infrastructure, personal devices, and edg

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

The Refusal Residue: When Probes Catch Alignment Faking and When They Don't

DGX agent

arXiv:2607.13346v1 Announce Type: cross Abstract: Alignment faking is dangerous because a model can appear compliant under monitoring while preserving behavior it would reveal when unmonitored. When n

model-releasesarxiv-cs-ai
16 Jul 2026
Tutorials

TreeSRNF: Square-Root Normal Fields for Generative Modelling of the Geometric and Structural Variability in Tree-like 3D Objects

DGX agent

arXiv:2607.13456v1 Announce Type: new Abstract: We introduce a novel mathematical framework for analyzing and generating complex tree-shaped 3D objects, such as botanical trees and plants, which defor

tutorialsarxiv-cs-cv
16 Jul 2026
Model Releases

Value Drifts: Tracing Value Alignment During LLM Post-Training

DGX agent

arXiv:2510.26707v2 Announce Type: replace-cross Abstract: As LLMs occupy an increasingly important role in society, they are more and more confronted with questions that require them not only to draw

model-releasesarxiv-cs-lg
16 Jul 2026
Model Releases

Accepted Prefixes Are Not All You Need: A Negative Result on PEFT-Based Block-Diffusion Drafting

DGX agent

arXiv:2607.12422v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive language model inference by using a cheap drafter to propose multiple future tokens and a target model t

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Do AI Agents Know When a Task Is Simple? Toward Complexity-Aware Reasoning and Execution

DGX agent

arXiv:2607.13034v1 Announce Type: new Abstract: Large language model (LLM) agents increasingly automate multi-step engineering and informatics workflows, yet they rarely ask how much effort a task act

model-releasesarxiv-cs-ai
15 Jul 2026
Safety

Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models

DGX agent

arXiv:2607.12463v1 Announce Type: new Abstract: Coding agents must integrate external tool returns into ongoing reasoning - a capability that standard left-to-right pretraining on code exposes only in

safetyarxiv-cs-ai
15 Jul 2026
Model Releases

GDP.pdf: Benchmarking Grounded Multimodal Reasoning over Professional PDF Documents

DGX agent

arXiv:2607.11192v2 Announce Type: replace Abstract: A large share of day-to-day work in professional domains happens inside PDF files: benefits packets, leases, datasheets, clinical guidelines, constr

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

Line-Anchored Feedback Cuts Token Costs and Improves Correctness in AI Code Editing

DGX agent

arXiv:2607.12713v1 Announce Type: cross Abstract: Generated tokens are a direct driver of the cost, latency, and energy of generative AI (GAI) code editing. We show the format of feedback is a lever o

model-releasesarxiv-cs-ai
15 Jul 2026
Safety

Predictive Modeling of High-Altitude Clear Air Turbulence in the United States: A Machine Learning Approach

DGX agent

arXiv:2607.11899v1 Announce Type: cross Abstract: High-altitude Clear Air Turbulence (CAT) poses significant risks to aviation safety due to its unpredictability and challenges in detection. This stud

safetyarxiv-cs-lg
15 Jul 2026
Research

Propheticus: Machine Learning Framework for the Development of Predictive Models for Reliable and Secure Software

DGX agent

arXiv:1809.01898v2 Announce Type: replace-cross Abstract: The growing complexity of software calls for innovative solutions that support the deployment of reliable and secure software. Machine Learnin

researcharxiv-cs-ai
15 Jul 2026
Model Releases

Reducing Temporal Redundancy for Efficient Vision-Language-Action Inference

DGX agent

arXiv:2607.12287v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models exhibit strong generalization for robotic manipulation, yet their high inference latency limits real time deployment

model-releasesarxiv-cs-ro
15 Jul 2026
Research

The Geometry of Memorization: Finite-Time Spectral Sensitivity as a Diagnostic for Flow Matching Models

DGX agent

arXiv:2607.12616v1 Announce Type: new Abstract: Continuous-time generative frameworks construct probability paths between base and target domains by optimizing time-dependent velocity fields. While th

researcharxiv-cs-lg
15 Jul 2026
Local Ai

Toward Localizing and Repairing Bias in Transformer Attention Heads

DGX agent

arXiv:2607.12863v1 Announce Type: cross Abstract: Transformer language models are increasingly used as software components, yet biased outputs remain difficult to localize and repair inside the model.

local-aiarxiv-cs-lg
15 Jul 2026
Agents

Traj-VLN: Learning Pixel-Space Interaction via Autoregressive Trajectory Generation

DGX agent

arXiv:2607.10744v2 Announce Type: replace Abstract: Benefiting from the powerful priors embedded in large-scale pre-training data and the emerging commonsense reasoning ability, large language models

agentsarxiv-cs-cv
15 Jul 2026
Model Releases

What Does a Temporal Benchmark Score Measure? Decomposing Channel Use in Video VLM Evaluation

DGX agent

arXiv:2607.12304v1 Announce Type: new Abstract: A score on a temporal video question answering benchmark is meant to measure that a model has temporal understanding, but it conflates two questions. 1.

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras

DGX agent

arXiv:2607.12993v1 Announce Type: new Abstract: We present X-lens, a compact feed-forward model for metric depth estimation from a variable number of calibrated fisheye and pinhole views. To support r

model-releasesarxiv-cs-cv
15 Jul 2026
Safety

EgoWAM: World Action Models Beyond Pixels with In-the-Wild Egocentric Human Data

DGX agent

arXiv:2607.08436v1 Announce Type: cross Abstract: Egocentric human data offers scalable supervision for robot manipulation. However, behavior cloning entangles transferable content like objects, scene

safetyarxiv-cs-ai
10 Jul 2026
Research

Enhancing the KidSat Model: Integrating Geographical Encoding and Data Quality Assessment for Childhood Poverty Prediction

DGX agent

arXiv:2607.08281v1 Announce Type: new Abstract: Accurate poverty mapping using satellite imagery is often hindered by (i) noisy and sparse survey-derived supervision, (ii) image quality issues such as

researcharxiv-cs-cv
10 Jul 2026
Tutorials

Frequency-Domain Multi-Modality Transportation Modeling

DGX agent

arXiv:2607.08475v1 Announce Type: new Abstract: Multi-modality transportation refers to urban systems composed of multiple transportation modes, such as traffic flow and public transit, whose dynamics

tutorialsarxiv-cs-lg
10 Jul 2026
Model Releases

Functional and Secure Code Generation with Task Vectors

DGX agent

arXiv:2607.07881v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for code generation, but they struggle to generate functional code free of security vulnerabilities

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

LEXIC: Lightweight Eye-tracking eXtension via Injected Complexity

DGX agent

arXiv:2607.08152v1 Announce Type: cross Abstract: On the recent EyeBench benchmark, predicting reading comprehension from eye movements exposes a stark gap: text-aware models using pretrained language

model-releasesarxiv-cs-ai
10 Jul 2026
Safety

LTM: Large-scale Terrain Model for Wildfire-prone Landscapes

DGX agent

arXiv:2607.08711v1 Announce Type: new Abstract: Accurate 3D terrain maps are essential for emergency response when assessing wildfire hazards. However, wildfire-prone regions often span vast areas whe

safetyarxiv-cs-cv
10 Jul 2026
Local Ai

PARA-PV: Physics-Aware Retrieval-Augmented PV Prediction Based on Frozen Foundation Model and Distribution Shift Correction

DGX agent

arXiv:2607.08079v1 Announce Type: new Abstract: Accurate photovoltaic (PV) power forecasting is essential for reliable grid dispatch and renewable energy integration, yet it remains challenging becaus

local-aiarxiv-cs-ai
10 Jul 2026
Applications

Reaction-network reasoning with frontier models for experimentally confirmed catalyst-selectivity hypotheses

DGX agent

arXiv:2607.08003v1 Announce Type: cross Abstract: Catalysts are essential for sustainable chemical manufacturing, yet discovering novel architectures remains a bottleneck dominated by trial-and-error

applicationsarxiv-cs-ai
10 Jul 2026
Safety

Search-based Testing of Vision Language Models for In-Car Scene Understanding

DGX agent

arXiv:2607.02300v2 Announce Type: replace Abstract: In the automotive domain, in-car scene understanding (ISU) enables the detection of safety-critical events, such as driver distraction, and supports

safetyarxiv-cs-cv
10 Jul 2026
Model Releases

Secure Decentralized Federated Learning via Gossip and Virtual Voting

DGX agent

arXiv:2607.08651v1 Announce Type: new Abstract: Decentralized federated learning (DFL) removes the central server by letting nodes exchange model updates through peer-to-peer gossip, but existing goss

model-releasesarxiv-cs-lg
10 Jul 2026
Applications

Unpaired Joint Distribution Modeling via Multi-Scale Image Representations

DGX agent

arXiv:2607.08198v1 Announce Type: new Abstract: This paper studies the problem of learning a joint distribution from marginal observations, which is inherently ill-posed due to the ambiguity of feasib

applicationsarxiv-cs-cv
10 Jul 2026
Research

Attention in Geometry: Scalable Spatial Modeling via Adaptive Density Fields and FAISS-Accelerated Kernels

DGX agent

arXiv:2601.06135v3 Announce Type: replace-cross Abstract: Spatial computation in geographic systems increasingly requires query-conditioned, local, interpretable aggregation under metric constraints.

researcharxiv-cs-cv
9 Jul 2026
Model Releases

Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages

DGX agent

arXiv:2607.06596v1 Announce Type: cross Abstract: Trusted monitoring is a central defense in AI control: a cheaper trusted model scores an untrusted model's actions for sabotage, and the most suspicio

model-releasesarxiv-cs-lg
9 Jul 2026
Tutorials

Generative Diffusion Models of Stochastic Graph Signals

DGX agent

arXiv:2607.06833v1 Announce Type: new Abstract: Sampling stochastic signals supported on a graph underlies many graph machine learning tasks, including recommender systems, forecasting in financial ma

tutorialsarxiv-cs-lg
9 Jul 2026
Model Releases

Geometric Self-Distillation for Reasoning Generalization

DGX agent

arXiv:2607.06855v1 Announce Type: cross Abstract: On-policy distillation is a practical post-training recipe for large language models, supplying dense teacher supervision on the student's own traject

model-releasesarxiv-cs-cl
9 Jul 2026
Research

On Adversarial Vulnerability of Vision-Language Models through the Lens of Intermediate Spectral Subspaces

DGX agent

arXiv:2607.07375v1 Announce Type: cross Abstract: Adversarial vulnerability in deep neural networks (DNNs) has been studied from the perspectives of decision-boundary geometry, feature robustness, inp

researcharxiv-cs-ai
9 Jul 2026
Model Releases

The Appeal and Reality of Recycling LoRAs with Adaptive Merging

DGX agent

arXiv:2602.12323v2 Announce Type: replace Abstract: The widespread availability of fine-tuned LoRA modules for open pre-trained models has led to an interest in methods that can adaptively merge LoRAs

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

The Harness Effect: How Orchestration Design Sets the Token Economics of Enterprise Agentic AI

DGX agent

arXiv:2607.06906v1 Announce Type: new Abstract: Agentic AI development today runs on token maxing: buying capability with tokens -- longer reasoning traces, more turns, wider tool payloads, bigger rep

model-releasesarxiv-cs-ai
9 Jul 2026
Research

Estimation of instrument and noise parameters for inverse problem based on prior diffusion model

DGX agent

arXiv:2602.11711v2 Announce Type: replace-cross Abstract: This article addresses the issue of estimating observation parameters (response and error parameters) in inverse problems. The focus is on cas

researcharxiv-cs-lg
8 Jul 2026
Agents

EvalLoop: A Methodology for Evaluation-Driven Iterative Improvement of Business AI Systems

DGX agent

arXiv:2607.05638v1 Announce Type: cross Abstract: Teams deploying large language models in business contexts need evaluation systems, yet most treat evaluation as static model selection: run benchmark

agentsarxiv-cs-ai
8 Jul 2026
Safety

From Blueprint to Reality: Modeling and Applying Putnam's Social Capital Theory with LLM-based Multi-agent Simulations

DGX agent

arXiv:2607.06080v1 Announce Type: cross Abstract: Putnam's Social Capital Theory is a foundational framework for collective action and community prosperity. However, traditional empirical methods face

safetyarxiv-cs-ai
8 Jul 2026
Research

i-EXAM: Instructable and Explainable Attack Connectivity Graph Modeler

DGX agent

arXiv:2607.05888v1 Announce Type: cross Abstract: i-EXAM is a planning-powered tool that helps system administrators to create security profiles of complex networks and perform what-if analyses to ide

researcharxiv-cs-ai
8 Jul 2026
Model Releases

Learning to Control LLM Agent Harnesses with Offline Reinforcement Learning

DGX agent

arXiv:2607.05458v1 Announce Type: cross Abstract: Large language model (LLM) agents are usually improved by changing prompts, models, or hand-written workflows, while the execution harness around the

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

LongCrafter: Towards Diverse Long-Context Understanding via Evidence-Graph-Guided Instruction Synthesis

DGX agent

arXiv:2607.06160v1 Announce Type: cross Abstract: Synthesizing long-context supervised fine-tuning (SFT) data is a scalable way to enhance the long-context understanding of large language models (LLMs

model-releasesarxiv-cs-ai
8 Jul 2026
← Previous
1…254255256257258…1058
Next →