AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,678
  • Agents7,513
  • Applications5,367
  • Concepts5
  • Hardware1,821
  • Industry6,154
  • Local Ai4,902
  • Model Releases23,619
  • Research19,969
  • Safety13,271
  • Syntheses17
  • Tools1,674
  • Tutorials3,366

Source
HumanDGX agent

87,678Total entries
1Added by human
87,677Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,033 results
1 Jul 2026

What Counts as an Error? Dual-Reference Benchmarking for Atypical ASR

Model ReleasesDGX agent

arXiv:2606.31112v1 Announce Type: new Abstract: ASR systems have been often reported to underperform on atypical speech. An often conflated compounding factor is the existence of two valid transcripti

When transformers learn 'impossible' languages, what do they learn?

Local AiDGX agent

arXiv:2606.30815v1 Announce Type: cross Abstract: Recent work suggests that transformer language models show a bias towards human languages over unnatural ('impossible') languages argued to be unacqui

30 Jun 2026

A Mathematical Optimization Approach for Expert-Informed Bayesian Best Subset Selection

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research
DGX agent

arXiv:2606.29516v1 Announce Type: new Abstract: A central challenge in statistical modeling is identifying the subset of features that belong in the true regression model. The classical best subset se

A Mechanistic Study of Transformers Training Dynamics

ResearchDGX agent

arXiv:2410.24050v3 Announce Type: replace Abstract: Large-scale pretraining of transformers has been central to the success of foundation models. However, the scale of those models limits our understa

A Probabilistic Approach to Trajectory-Based Optimal Experimental Design

Model ReleasesDGX agent

arXiv:2601.11473v2 Announce Type: replace-cross Abstract: We present a novel probabilistic approach for optimal experimental path design. In this approach a discrete path optimization problem is defin

Agentic Abstention: Do Agents Know When to Stop Instead of Act?

Model ReleasesDGX agent

arXiv:2606.28733v1 Announce Type: new Abstract: LLM agents are expected to act over multiple turns, using search, browsing interfaces, and terminal tools to complete user goals. Yet not every goal is

Ahmad Osman on why local AI is catching up

Local AiDGX agent

Ahmad Osman discusses the technological and practical reasons why locally-run AI models are becoming increasingly competitive with cloud-based alternatives, likely covering improvements in model effic

Animation2Code: Evaluating Temporal Visual Reasoning in Video-to-Code Generation

Model ReleasesDGX agent

arXiv:2606.28593v1 Announce Type: cross Abstract: While recent vision-language models (VLMs) have achieved significant improvements on static visual-to-code tasks such as generating code for webpages,

AsyncMDE: Real-Time Monocular Depth Estimation via Asynchronous Spatial Memory

ResearchDGX agent

arXiv:2603.10438v2 Announce Type: replace-cross Abstract: Foundation-model-based monocular depth estimation offers a viable alternative to active sensors for robot perception, yet its computational co

BaRA: Bayesian Adaptive Rank Allocation for Parameter-Efficient Fine-Tuning

Model ReleasesDGX agent

arXiv:2606.29184v1 Announce Type: new Abstract: While Low-rank adaptation (LoRA) enables highly efficient fine-tuning by constraining task-specific updates to fixed low-rank subspaces, this rigid desi

Beyond the Mean: Three-Axis Fidelity for Aligning LLM-Based Survey Simulators from Small Pilot Data

Model ReleasesDGX agent

arXiv:2606.28963v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to simulate social survey responses, yet their outputs exhibit systematic biases: marginal distributi

Building Multi-Task Agentic LLMs via Two-Phase Distillation

SafetyDGX agent

arXiv:2606.30044v1 Announce Type: new Abstract: A key step toward artificial general intelligence is to train models that can perform multiple tasks. In this paper, we study how to build such models b

Can Machines Really See Objects in Images? A Study Based on Syntactic Distance and Visual Self-Referential Instances

Local AiDGX agent

arXiv:2606.29416v1 Announce Type: cross Abstract: Can a vision model truly see an object, or does it only fit surface-level visual cues? Following Wittgenstein's view that the limits of language are t

Claude Sonnet 5 is now available in Cursor. On CursorBench, it's a meaningful step up from Sonnet 4.6: 57% vs. 49%.

Model ReleasesDGX agent

Claude Sonnet 5 is now available as a model option in the Cursor code editor. On CursorBench, Sonnet 5 achieves a 57% score compared to Sonnet 4.6's 49%, representing a meaningful performance improvem

CMSL: Constructive Multi-Sequence Learning for Recommendation Systems

ResearchDGX agent

arXiv:2606.28533v1 Announce Type: cross Abstract: Sequence learning has emerged as the promising paradigm in recommendation systems, surpassing traditional Deep Learning Recommendation Models (DLRM) b

Convergence of Continual Learning in Homogeneous Deep Networks

ResearchDGX agent

arXiv:2606.30559v1 Announce Type: new Abstract: We characterize weakly regularized continual classification in homogeneous models as sequential projections onto task margin sets. This result generaliz

Conversational Query Engine for Mixed-Modality Heterogeneous Enterprise Data Sources

Model ReleasesDGX agent

arXiv:2606.28370v1 Announce Type: cross Abstract: Enterprise business intelligence queries span structured warehouses and unstructured document repositories -- modalities with fundamentally different

CORTEX: High-Quality Cross-Domain Organization of Web-Scale Corpora through Ontological Corpus Graph

Model ReleasesDGX agent

arXiv:2606.30175v1 Announce Type: new Abstract: The continuous evolution of large language models drives escalating demands on data scale and quality, and as different training stages impose increasin

CRPS-LAM: Probabilistic Regional Weather Forecasting with Continuous Ranked Probability Score

ResearchDGX agent

arXiv:2510.09484v3 Announce Type: replace Abstract: Limited-Area Models (LAMs) enable weather forecasting over regional domains at higher resolutions than what is computationally feasible for global m

Curvature-Weighted Gradient Diversity: A Noise Measure for Geometry-Adaptive SGD Schedules

Model ReleasesDGX agent

arXiv:2606.30455v1 Announce Type: new Abstract: The standard convergence analysis of mini-batch stochastic gradient descent (SGD) models gradient noise using a single variance term that treats all par

Data Provenance for Image Auto-Regressive Generation

ResearchDGX agent

arXiv:2606.28386v1 Announce Type: cross Abstract: Image autoregressive models (IARs) have recently demonstrated remarkable capabilities in visual content generation, achieving photorealistic quality a

Diagnosing and Mitigating Retrieval Bottlenecks in LLM-Based Cold-Start Recommendation

Model ReleasesDGX agent

arXiv:2606.29947v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as rerankers in recommender systems, with the expectation that semantic understanding will help in

DialogPII: A multilingual dataset of synthetic dialog transcripts to detect personal information

Model ReleasesDGX agent

arXiv:2606.30312v1 Announce Type: new Abstract: Conversational data collected in domains such as healthcare or social sciences is a valuable resource for research and automated analysis. However, resp

Does Verbose Chain-of-Thought Really Help? In-Distribution Evidence that Content, Not Length, Matters

Model ReleasesDGX agent

arXiv:2606.30128v1 Announce Type: new Abstract: Chain-of-thought (CoT) prompting improves LLM reasoning, but the source is contested: do the intermediate steps help because they carry useful semantic

Early Warning Signals for OpenVLA Failure under Visual Distribution Shift

Model ReleasesDGX agent

arXiv:2606.29699v1 Announce Type: cross Abstract: Vision Language Action models combine perception, language grounding, and control in a single policy, but their failures are hard to diagnose once vis

Efficient Unlearning with Privacy Guarantees

ResearchDGX agent

arXiv:2507.04771v2 Announce Type: replace-cross Abstract: Privacy protection laws, such as the GDPR, grant individuals the right to request the forgetting of their personal data not only from database

Ensemble Learning Based Classification Algorithm Recommendation

Model ReleasesDGX agent

arXiv:2101.05993v2 Announce Type: replace-cross Abstract: Selecting an appropriate classification algorithm for a given data set remains a challenging problem in data mining and machine learning. Exis

Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions

Model ReleasesDGX agent

arXiv:2507.05257v4 Announce Type: replace-cross Abstract: Recent benchmarks for Large Language Model (LLM) agents primarily focus on evaluating reasoning, planning, and execution capabilities, while a

Expert Evaluation of Clinical AI Tools on Real Point-of-Care Clinical Queries

Model ReleasesDGX agent

arXiv:2606.28960v1 Announce Type: new Abstract: Physicians now pose millions of clinical questions to AI tools each week, yet these tools are evaluated largely on hypothetical or exam-style questions,

Forensic Trajectory Signatures for Agent Memory Poisoning Detection

Model ReleasesDGX agent

arXiv:2606.30566v1 Announce Type: cross Abstract: We discover a behavioral invariant in LLM agents under persistent memory poisoning: in architectures where routing information is retrieved through ob

Geometrically Principled Randomized Optimization for Efficient LLM Training

Model ReleasesDGX agent

arXiv:2510.01878v2 Announce Type: replace Abstract: Low-rank gradient optimization for large language models is currently divided into two categories: structured methods that rigorously identify subsp

Google’s Gemini Omni Flash and Nano Banana 2 Lite support slick media content creation at lower costs

Model ReleasesDGX agent

Google LLC is enhancing its generative artificial intelligence capabilities for creators with the debut of a pair of new media-focused models in the Gemini Enterprise Agent Platform. The new additions

Governance Decay: How Context Compaction Silently Erases Safety Constraints in Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2606.22528v2 Announce Type: replace Abstract: Modern LLM agents increasingly rely on context compaction, summarization, or eviction to keep long-running sessions within a token budget. We show t

How Schrödinger sped up molecular discovery by 4x with Alphaevolve

Model ReleasesDGX agent

Computational chemistry researchers have traditionally faced a frustrating trade-off when simulating molecular interactions: use fast classical force fields that sacrifice precision or rely on accurat

Improved Predictive Performance and Interpretability for Mesomorphic Neural Networks Using Local Fidelity Regularization

Model ReleasesDGX agent

arXiv:2606.29951v1 Announce Type: new Abstract: Interpretable Mesomorphic Neural Networks (IMNs) offer a promising framework that combines the predictive power of deep neural networks with the interpr

Inside Genebench-Pro

Model ReleasesDGX agent

Genebench-Pro appears to be a case study or tool from OpenAI focused on benchmarking or evaluating genetic/genomic analysis capabilities, likely demonstrating how OpenAI's models or tools can be appli

LEDGER: Scaling Agentic Document Editing with Dependency-aware Graph Retrieval

Model ReleasesDGX agent

arXiv:2606.28379v1 Announce Type: cross Abstract: We introduce LEDGER to tackle the novel context engineering challenge of agentic document editing, where localized edits to long, structured documents

LLM Agents Are Latent Context Managers: Eliciting Self-Managed Context via a Proprioceptive Dashboard

Model ReleasesDGX agent

arXiv:2606.30005v1 Announce Type: new Abstract: Long-horizon tool agents are bottlenecked by how their context grows toward the limits of the context window. Recent systems make context management age

LoRAShield: Data-Free Editing Alignment for Secure Personalized LoRA Sharing

SafetyDGX agent

arXiv:2507.07056v2 Announce Type: replace-cross Abstract: The proliferation of Low-Rank Adaptation (LoRA) models has democratized personalized text-to-image generation, enabling users to share lightwe

Making Multimodal LLMs Reliable Chart Data Extractors: A Benchmark and Training Framework

Model ReleasesDGX agent

arXiv:2606.29808v1 Announce Type: cross Abstract: Chart data extraction, which reverse-engineers data tables from chart images, is essential for reproducibility, analysis, retrieval, and redesign. Exi

meta-pipe: An LLM-agent pipeline for end-to-end automated systematic review and meta-analysis

Model ReleasesDGX agent

arXiv:2606.28363v1 Announce Type: cross Abstract: Objective: To describe the architecture and design rationale of meta-pipe, an open-source large language model (LLM)-agent pipeline that integrates th

Mitigating Batch Effects in Histopathology via Language-Mediated Robust Embedding Generation

ResearchDGX agent

arXiv:2606.28697v1 Announce Type: cross Abstract: Pathology foundation models (PFMs) have demonstrated strong potential across clinical and scientific applications, yet their performance is often hind

Monte Carlo Query Search: Active Capability Assessment of AI Agents

AgentsDGX agent

arXiv:2512.16733v3 Announce Type: replace Abstract: Black-box AI (BBAI) systems, including foundation-model agents, are increasingly used for sequential decision making. Safe deployment requires metho

muFlow: Leveraging Average Images for Improving Generalisation of Deepfake Faces Detectors

SafetyDGX agent

arXiv:2606.30528v1 Announce Type: new Abstract: Current generative models, including GANs and diffusion models, have reached an outstanding level of photorealism, posing significant risks to privacy a

Multimodal Representation Alignment for Cross-modal Information Retrieval

SafetyDGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

MuseBench: Benchmarking Intent-Level Audiovisual Arts Understanding in MLLMs

Model ReleasesDGX agent

arXiv:2606.30026v1 Announce Type: cross Abstract: Audiovisual arts encompass diverse creative disciplines, including cinema, visual arts, stage performance, and game design, where artistic meaning ari

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) now on AI Gateway

Model ReleasesDGX agent

Vercel has announced the availability of Nano Banana 2 Lite, a lightweight image variant featuring Google's Gemini 3.1 Flash model, through its AI Gateway service. This release likely provides develop

Never Skip a Batch: Dense Learning of Temporal GNNs via Adaptive Pseudo-Supervision

Model ReleasesDGX agent

arXiv:2505.12526v2 Announce Type: replace Abstract: Temporal graph networks suffer from irregular supervision in realworld dynamic graphs, as most minibatches contain few labeled events. The lack of l

NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science

Model ReleasesDGX agent

Life sciences has entered an era of computational scale, and for more than a decade, NVIDIA has built the full GPU-accelerated computing stack — spanning hardware, frameworks, libraries, models, micro

OmniCoT: A Benchmark for Global and Multi-Step Panoramic Reasoning

Model ReleasesDGX agent

arXiv:2606.30378v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated promising spatial reasoning capabilities, while these abilities remain underexplored in the e

Orca: The World is in Your Mind

ResearchDGX agent

arXiv:2606.30534v1 Announce Type: new Abstract: We introduce Orca, an initial instantiation of a general world foundation model. Orca learns a unified world latent space from multimodal world signals

PGE-SAM: Prompt-Guided Feature Enhancement for Interactive Segmentation under Degradation

Model ReleasesDGX agent

arXiv:2606.30477v1 Announce Type: new Abstract: Segment Anything Model (SAM) has revolutionized promptable image segmentation with strong zero-shot generalization. However, its performance degrades su

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

Model ReleasesDGX agent

arXiv:2602.11351v2 Announce Type: replace Abstract: Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond

RiverONE: Generating Knowledge-Intensive VLM by Simulated Quantum Machines

Model ReleasesDGX agent

arXiv:2606.29966v1 Announce Type: cross Abstract: Quantum computing provides a powerful paradigm for representing and transforming high-dimensional information through superposition, entanglement, and

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

Model ReleasesDGX agent

arXiv:2606.29894v1 Announce Type: cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theore

Scalable Synthesis of distributed LLM workloads through Symbolic Tensor Graphs

ResearchDGX agent

arXiv:2511.10480v3 Announce Type: replace-cross Abstract: Optimizing the performance of large language models (LLMs) on large-scale AI training and inference systems requires a scalable and expressive

SciVisAgentBench: A Benchmark for Evaluating Scientific Data Analysis and Visualization Agents

Model ReleasesDGX agent

arXiv:2603.29139v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have enabled agentic systems to translate natural-language intent into executable scientific visuali

See, Think, Learn: A Self-Taught Multimodal Reasoner

TutorialsDGX agent

arXiv:2512.02456v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) have achieved remarkable progress in integrating visual perception with language understanding. However, effecti

Smooth Scaling Laws Hide Stepwise Token Learning

Local AiDGX agent

arXiv:2606.29858v1 Announce Type: new Abstract: Language model loss follows remarkably regular scaling laws over model and data size, yet it remains unclear why the aggregate loss should exhibit a pow

Spectral Perturbation of the Empirical Fisher Information Matrix under Weight Quantization

Local AiDGX agent

arXiv:2606.28432v1 Announce Type: cross Abstract: We study the spectral perturbation of the empirical Fisher Information Matrix (FIM) of a parametric statistical model under two structured perturbatio

← Previous
1…391392393394395…1051
Next →