AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

Reproducible Reservoir Computing with Thermally Driven Superparamagnets: Controlling Temperature Sensitivity

DGX agent

arXiv:2607.12840v1 Announce Type: cross Abstract: Unconventional computing systems must demonstrate robust performance under real-world environmental conditions to enable practical deployments. We hav

model-releasesarxiv-cs-ai
15 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Resist and Update: Counterfactual Report Coordinates for Incentive-Compatible LLMs

DGX agent

arXiv:2607.12985v1 Announce Type: new Abstract: Aligned language models routinely misreport under non-evidential incentive pressure: they agree with a confident user or overstate certainty even when t

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Rethinking the Evaluation of Harness Evolution for Agents

DGX agent

arXiv:2607.12227v1 Announce Type: new Abstract: We revisit the evaluation of automatic harness evolution for LLM agents. Existing harness evolution methods use unit test cases to search for harness co

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

RippleBench: Capturing Ripple Effects Using Existing Knowledge Repositories

DGX agent

arXiv:2512.04144v3 Announce Type: replace Abstract: Targeted interventions on language models, such as unlearning or model editing, aim to modify specific information, but their effects often propagat

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

RoboDesign1M: A Large-scale Dataset for Robot Design Understanding

DGX agent

arXiv:2503.06796v2 Announce Type: replace Abstract: Robot design is a complex and time-consuming process that requires specialized expertise. Gaining a deeper understanding of robot design data can en

model-releasesarxiv-cs-ro
15 Jul 2026
Model Releases

Sample Efficient Generative Optimization for Molecular Design

DGX agent

arXiv:2607.12488v1 Announce Type: new Abstract: Molecular optimization in drug discovery, materials design, and catalysis requires searching vast chemical spaces under tight evaluation budgets, since

model-releasesarxiv-cs-lg
15 Jul 2026
Model Releases

Scale-Aware Attention for Scarce Neural Data: An RG-Flow Transformer on Sleep-EDF EEG

DGX agent

arXiv:2607.11950v1 Announce Type: cross Abstract: Brain field potentials are scale-free: their power spectra follow a 1/f^{eta} law whose aperiodic exponent eta tracks cortical state, and sleep depth

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Scaling Point-in-Time Language Models

DGX agent

arXiv:2607.11889v1 Announce Type: cross Abstract: Large language models trained on unrestricted internet corpora inevitably embed information from the future, introducing lookahead bias that compromis

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Self-Evolving In-Context Learning for Direct Pilot-to-Beamformer Design in MU-MISO Systems

DGX agent

arXiv:2607.11970v1 Announce Type: cross Abstract: We develop an enhanced in-context learning (ICL) framework to improve the performance of pilot-based beamforming in multi-user multiple-input single-o

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Self in Space: Benchmarking Self-Awareness and Spatial Cognition in UAV Embodied Intelligence

DGX agent

arXiv:2607.12477v1 Announce Type: new Abstract: Autonomous UAV systems increasingly rely on multimodal large language models (MLLMs) to operate in complex real-world environments. Such embodied scenar

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

Sensitivity to Subjective Expected Utility Maximization: A Methodological Study, with an Illustrative Application to LLM Decision-Making

DGX agent

arXiv:2607.11920v1 Announce Type: cross Abstract: Evaluating decisions made under uncertainty is hard when labeled outcomes are scarce, costly, or confounded with luck. We treat subjective expected ut

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

SeqGPT: A Constrained Transformer Agent for the Inverse Designof Multi-Panel Composite Structures

DGX agent

arXiv:2607.11910v1 Announce Type: cross Abstract: Optimizing composite stacking sequences to match continuous targets (e.g., Lamination or Buckling Parameters) with discrete manufacturing constraints

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

SheetMind: An End-to-End LLM-Powered Multi-Agent Framework for Spreadsheet Automation

DGX agent

arXiv:2506.12339v2 Announce Type: replace-cross Abstract: We present SheetMind, a modular multi-agent framework powered by large language models (LLMs) for spreadsheet automation via natural language

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Silent Alarm: A J-Space Protocol for Comparing Danger Recognition Across Models and Quantization Levels

DGX agent

arXiv:2607.12792v1 Announce Type: cross Abstract: Jailbreak-robustness research typically evaluates safety through generated responses using an LLM-as-judge approach. Such evaluations, however, are se

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

SKooP: Symmetric Koopman Predictions for Faster and More Generalizable Legged Robot Locomotion with Reinforcement Learning

DGX agent

arXiv:2607.11624v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) algorithms classically suffer from poor sample efficiency. In robotics, a recent line of work has emerged addressi

model-releasesarxiv-cs-lg
15 Jul 2026
Model Releases

So Many Opinions, So Many LLMs: Comparing Large Language Models to Traditional Machine Learning for Open- Ended Survey Analysis

DGX agent

arXiv:2607.11890v1 Announce Type: cross Abstract: Open-ended surveys offer valuable insights, but they are notoriously difficult to analyze at scale. Building on previous work that employed traditiona

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Spherical-GOF: Geometry-Aware Panoramic Gaussian Opacity Fields for 3D Scene Reconstruction

DGX agent

arXiv:2603.08503v2 Announce Type: replace Abstract: Omnidirectional images are increasingly used in robotics and vision due to their wide field of view. However, extending 3D Gaussian Splatting (3DGS)

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise

DGX agent

arXiv:2602.12783v3 Announce Type: replace-cross Abstract: Spoken query retrieval is an important interaction mode in modern information retrieval. However, existing evaluation datasets are often limit

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

TerraLogic: A Benchmark for Hierarchical Geospatial Reasoning in Earth Observation

DGX agent

arXiv:2607.12497v1 Announce Type: new Abstract: Beyond perception, reasoning is essential in remote sensing for advanced interpretation, inference, and decision-making. Recent advances in large langua

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

TerraZero: Procedural Driving Simulation for Zero-Demonstration Self-Play at Scale

DGX agent

arXiv:2607.13028v1 Announce Type: cross Abstract: Training robust autonomous driving agents requires a simulator that is fast enough for reinforcement learning at scale, realistic enough to ground beh

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

The Benjamini--Hochberg Procedure Can Fail to Control the FDR for Correlated Two-Sided Gaussian Tests

DGX agent

arXiv:2607.12208v1 Announce Type: cross Abstract: We show that the Benjamini--Hochberg procedure can fail to control the false discovery rate (FDR) at its nominal level for correlated two-sided Gaussi

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

The Capacity of Thought: Benchmarking Llama 3.2 in Semantic fMRI Neural Language Decoding and Improving the Huth Encoding-Model Baseline

DGX agent

arXiv:2607.12079v1 Announce Type: new Abstract: Decoding continuous language from fMRI signals remains a core challenge in non-invasive brain-computer interface research. We present two complementary

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

The Illusion of Robustness: Aggregate Accuracy Hides Prediction Flips under Task-Irrelevant Context

DGX agent

arXiv:2607.12963v1 Announce Type: new Abstract: As large language models (LLMs) grow more capable, they are increasingly deployed in context-rich settings where task inputs are often accompanied by lo

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

The One-Word Census: Answer-Choice Conformity Across 44 Language Models

DGX agent

arXiv:2607.12796v1 Announce Type: cross Abstract: When a language model must pick one answer from a large space of equally valid options, which does it pick -- and how often is it the same answer ever

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

The TopCoW Challenge -- Topology-Aware Circle of Willis Segmentation for CT and MR Angiography

DGX agent

arXiv:2312.17670v5 Announce Type: replace Abstract: The Circle of Willis (CoW) is an important network of arteries connecting major circulations of the brain. Its vascular architecture is believed to

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

Token Reduction Is Not Cost Reduction

DGX agent

arXiv:2607.12161v1 Announce Type: new Abstract: Context-reduction layers for API-based coding agents, including command-output compressors, retrieval rankers, and payload-optimizing proxies, are usual

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Towards Self-Evolving Agents: A Human-Inspired Adaptive Exploration-Exploitation Framework for Genetic Network Programming

DGX agent

arXiv:2607.11913v1 Announce Type: cross Abstract: Recent advancements in agentic AI have increasingly moved toward graph-based methods, driven by the demand for explainable, human-centered, and non-li

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

TRACE: An Operational Reasoning Schema for Auditable Agentic Commitments

DGX agent

arXiv:2607.12480v1 Announce Type: new Abstract: This paper defines TRACE (Typed Reasoning And Commitment Evidence): a typed, versioned schema for recording reasoning traces, a reference procedure for

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Track, Rank, Crack: Epistemic Working Memory Scales Multi-Hop Reasoning in Language Agents

DGX agent

arXiv:2607.12267v1 Announce Type: cross Abstract: Language agents that interleave reasoning and tool use degrade sharply as reasoning chains lengthen, even when each individual step is easy. We trace

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

Transforming LLMs into Efficient Cross-Encoders via Knowledge Distillation for RAG Reranking

DGX agent

arXiv:2607.11933v1 Announce Type: new Abstract: Cross-encoders achieve high reranking accuracy in Retrieval-Augmented Generation (RAG) pipelines but impose quadratic inference costs that limit real-ti

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

UniVR: Thinking in Visual Space for Unified Visual Reasoning

DGX agent

arXiv:2607.12800v1 Announce Type: new Abstract: Learning broad world knowledge directly from raw visual data is a fundamental capability of intelligence. We introduce UniVR, the first investigation in

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

VanillaBench: The Hidden Accuracy Cost of Adversarial Robustness

DGX agent

arXiv:2607.12545v1 Announce Type: cross Abstract: Adversarial robustness research has produced hundreds of defended models over the past decade, yet the literature almost universally reports robustnes

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

Verifier-Based Reinforcement Fine-Tuning of Reasoning Models for Thermal Energy Storage Control

DGX agent

arXiv:2607.12856v1 Announce Type: new Abstract: Buildings are expected to shift cooling loads in response to grid conditions. Thermal energy storage (TES) enables this shift, but scheduling it well re

model-releasesarxiv-cs-lg
15 Jul 2026
Model Releases

ViHoRec: A Quality-Controlled Vietnamese Hotel Recommendation Dataset and Cold-Start Benchmark

DGX agent

arXiv:2607.12946v1 Announce Type: cross Abstract: Recommender-system research for Vietnamese remains limited by the absence of a public, well-documented hotel interaction resource. Building such a res

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

VisCo: Leveraging Large Language Models as Intrinsic Encoders for Visual Token Compression

DGX agent

arXiv:2607.12756v1 Announce Type: new Abstract: Vision-language models (VLMs) process large numbers of visual tokens, resulting in substantial inference latency and memory overhead. This has motivated

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

What Does a Temporal Benchmark Score Measure? Decomposing Channel Use in Video VLM Evaluation

DGX agent

arXiv:2607.12304v1 Announce Type: new Abstract: A score on a temporal video question answering benchmark is meant to measure that a model has temporal understanding, but it conflates two questions. 1.

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

When Close Enough Is Not Enough: Autoregressive Drift in Quantum Circuit Synthesis

DGX agent

arXiv:2607.12780v1 Announce Type: cross Abstract: Quantum circuit optimization for fault-tolerant computing requires exact functional equivalence while minimizing expensive non-Clifford resources such

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

When Directional Accuracy Lies: A Base-Rate-Honest Benchmark for LoRA-Adapted TimesFM on Equity Forecasting

DGX agent

arXiv:2607.12248v1 Announce Type: cross Abstract: Large pretrained time-series models such as TimesFM are attractive for financial forecasting, but raw directional accuracy is a misleading scoreboard

model-releasesarxiv-cs-lg
15 Jul 2026
Model Releases

WikiSTAR: A System for Shedding Light on the Hidden History of Scientific Wikipedia Articles

DGX agent

arXiv:2607.12441v1 Announce Type: new Abstract: Wikipedia plays a key role in shaping public understanding of science, and its openly accessible revision history is a unique record of how scientific k

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

X-Lens: Real-Time Metric Depth Estimation with Heterogeneous Cameras

DGX agent

arXiv:2607.12993v1 Announce Type: new Abstract: We present X-lens, a compact feed-forward model for metric depth estimation from a variable number of calibrated fisheye and pinhole views. To support r

model-releasesarxiv-cs-cv
15 Jul 2026
Model Releases

A law of robustness for two-layer neural networks with arbitrary weights

DGX agent

arXiv:2607.07778v1 Announce Type: new Abstract: Bubeck, Li and Nagaraj conjectured that, for generic data, any two-layer neural network with m neurons that fits n noisy labels must have Lipschitz cons

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

A Practical Investigation of Training-free Relaxed Speculative Decoding

DGX agent

arXiv:2607.08690v1 Announce Type: cross Abstract: Speculative decoding accelerates sampling from an autoregressive LLM by using a faster auxiliary model to draft tokens which are then verified in para

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

A Reliability Assessment of LALM Audio Judges for Full-Duplex Voice Agents

DGX agent

arXiv:2607.07985v1 Announce Type: cross Abstract: We report the empirical reliability of Gemini models as audio judges that score full-duplex agent conversations directly from the raw stereo waveform,

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

A safety-oriented hypothetico-deductive framework for AI-assisted differential diagnosis

DGX agent

arXiv:2607.08038v1 Announce Type: new Abstract: Diagnostic error is a major threat to patient safety, yet current large language model (LLM) systems often treat diagnosis as a one-shot prediction task

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

A Vision Toward Energy-Efficient Domain-Specific Artificial Intelligence Models and Agents

DGX agent

arXiv:2510.22052v2 Announce Type: replace Abstract: The field of artificial intelligence (AI) has taken a tight hold on broad aspects of society, industry, business, and governance in ways that dictat

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Adaptive Generation of Bias-Eliciting Questions for LLMs

DGX agent

arXiv:2510.12857v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now widely deployed in user-facing applications, reaching hundreds of millions of users worldwide. Despite th

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning

DGX agent

arXiv:2607.07761v1 Announce Type: new Abstract: Large language models (LLMs) have emerged as important tools in healthcare, showing growing potential for clinical reasoning and patient care. This surv

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Answer Set Programming Energised! End-to-End Neurosymbolic Reasoning and Learning with ASP and Energy Based Models

DGX agent

arXiv:2607.08136v1 Announce Type: new Abstract: We present a general neurosymbolic reasoning and learning methodology based on a modular integration of answer set programming with an energy based mode

model-releasesarxiv-cs-ai
10 Jul 2026
← Previous
1…7475767778…361
Next →