AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Model Releases

Scaling Trends for Lie Detector Oversight in Preference Learning

DGX agent

arXiv:2607.01567v1 Announce Type: new Abstract: Deceptive behavior in LLMs is costly to monitor and prevent, motivating approaches such as Scalable Oversight via Lie Detectors (SOLiD) (Cundy & Gleave,

model-releasesarxiv-cs-ai
3 Jul 2026
Local Ai
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Self-explainable Operator Learning for Discovering Spatial Patterns in Functional Data

DGX agent

arXiv:2607.02203v1 Announce Type: new Abstract: Operator learning has emerged as a powerful tool for modeling complex physical systems in functional spaces. However, their neural network-based archite

local-aiarxiv-cs-lg
3 Jul 2026
Model Releases

Separating Expert Retention from Autonomous Source Inference in Raw-ECG-Replay-Free Continual ECG Deployment

DGX agent

arXiv:2607.01674v1 Announce Type: new Abstract: In multi-source ECG deployment, models may need to incorporate new data sources when earlier raw ECGs cannot be retained or replayed. Freezing a pretrai

model-releasesarxiv-cs-ai
3 Jul 2026
Research

SPARCLE: SPeaker-aware Aligned Representations via Contrastive Language Embeddings

DGX agent

arXiv:2607.01238v1 Announce Type: cross Abstract: Recent advances in speech synthesis have shifted from phoneme representations to direct grapheme modeling. While phonemes address the one-to-many mapp

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Spec-AUF: Accept-Until-Fail Training under Train-Inference Misalignment for Masked Block Drafters

DGX agent

arXiv:2607.01893v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by drafting a block of tokens that the target model verifies left-to-right, committing only t

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

SPLIT: Cross-Lingual Empathy and Cultural Grounding in English and Ukrainian LLM Responses

DGX agent

arXiv:2607.02049v1 Announce Type: cross Abstract: Large Language Models are increasingly deployed in emotional-support contexts and crisis-related situations. Nevertheless, their cross-lingual abiliti

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Token Geometry

DGX agent

arXiv:2607.01455v1 Announce Type: cross Abstract: Language models learn continuous programs over discrete symbols, with the embedding table and LM-head acting as the read/write interface between them.

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

UA-ChatDev: Uncertainty-Aware Multi-Agent Collaboration for Reliable Software Development

DGX agent

arXiv:2607.02186v1 Announce Type: new Abstract: Software development is a complex task that demands cooperation among agents with diverse roles. Large language models (LLMs) have enabled autonomous mu

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

A Filtered Mixture-of-Generators for Fully Synthetic Survival Training

DGX agent

arXiv:2607.00127v1 Announce Type: new Abstract: Survival analysis models time-to-event data, but in clinical settings training data are costly and scarce: events accrue over years of follow-up, cohort

safetyarxiv-cs-lg
2 Jul 2026
Model Releases

AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation

DGX agent

arXiv:2607.00052v1 Announce Type: cross Abstract: GraphRAG is an extension of retrieval-augmented generation (RAG) that supports large language models (LLMs) by referring to graph-structured data as e

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

An LLM-Based Framework for Intent-Driven Network Topology Design

DGX agent

arXiv:2607.00292v1 Announce Type: cross Abstract: Designing deployable and resilient network topologies from natural language requirements remains a challenging problem in network automation. This wor

model-releasesarxiv-cs-ai
2 Jul 2026
Local Ai

Beyond Perplexity: A Behavioral Evaluation Framework for Deployment-Memory Claims in LLM Test-Time Training

DGX agent

arXiv:2607.00368v1 Announce Type: new Abstract: Large language model test-time training (TTT) is often evaluated through local proxy metrics: models are updated on recent tokens, retrieved context, ta

local-aiarxiv-cs-cl
2 Jul 2026
Model Releases

Can Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Use

DGX agent

arXiv:2607.01084v1 Announce Type: new Abstract: While Large Language Model (LLM) agents demonstrate proficiency in static benchmarks, their deployment in real-world scenarios is hindered by the dynami

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as w…

DGX agent

Continual learning is probably the biggest barrier to explosive AI adoption (& may have big implications for recursive self-improvement as well) As long as you deal with amnesiac models that require h

model-releasesethan-mollick--x
2 Jul 2026
Model Releases

Does Your ViT Still Need U-Net for Segmentation?

DGX agent

arXiv:2607.00223v1 Announce Type: new Abstract: Medical image segmentation is dominated by U-Net-style encoder-decoder architectures. Vision Transformers (ViTs) overcome the limited receptive field of

model-releasesarxiv-cs-cv
2 Jul 2026
Research

Flow-Map GRPO: Reinforcement Learning for Few-Step Flow-Map Generators via Anchored Stochastic Composition

DGX agent

arXiv:2607.00535v1 Announce Type: cross Abstract: Few-step flow-map generators, such as consistency models and MeanFlow, accelerate sampling by directly learning long-range transport maps between nois

researcharxiv-cs-ai
2 Jul 2026
Model Releases

FLYNN: Robust Neural Network for Robot Navigation using Fly Brain Topology

DGX agent

arXiv:2607.00025v1 Announce Type: cross Abstract: While deep learning models achieve state-of-the-art performance in complex tasks, they remain brittle when faced with new environments or sensory depr

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

GLM 5.2 DSpark preview is here! ✨ https://huggingface.co/RedHatAI/GLM-5.2-speculator.dspark-preview This is the first DSpark speculator for …

DGX agent

GLM 5.2 DSpark preview is here! ✨ https://huggingface.co/RedHatAI/GLM-5.2-speculator.dspark-preview This is the first DSpark speculator for a non-DeepSeek frontier model, trained with Speculators and

model-releasesclem-delangue--x
2 Jul 2026
Model Releases

GMO-E^2DIT: Grounded Multi-Operation Editing for E-Commerce Images

DGX agent

arXiv:2607.00920v1 Announce Type: new Abstract: Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

GPTKB v1.5: A Massive Knowledge Base for Exploring Factual LLM Knowledge

DGX agent

arXiv:2507.05740v2 Announce Type: replace Abstract: Language models are powerful artifacts, yet their factual knowledge is still poorly understood, and inaccessible to ad-hoc browsing and scalable sta

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

GSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV Cache

DGX agent

arXiv:2607.01065v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) with extended context windows is increasingly constrained by the linear growth of Key-Value (KV) cache me

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Homogenization of ell_2-Adversarial Training in High-Dimensions: Exact Dynamics under Stochastic Gradient Descent

DGX agent

arXiv:2607.00207v1 Announce Type: cross Abstract: We develop a framework for analyzing the learning dynamics of ell_2-adversarial training of single-index models on Gaussian mixtures in the high-dimen

model-releasesarxiv-cs-lg
2 Jul 2026
Research

Leveraging Multimodality for Real-Time Classification of Transients and Variables found by the Zwicky Transient Facility

DGX agent

arXiv:2607.00228v1 Announce Type: cross Abstract: Modern time-domain surveys such as the Zwicky Transient Facility (ZTF) generate hundreds of thousands of alerts each night, making real-time decisions

researcharxiv-cs-lg
2 Jul 2026
Model Releases

Lost in the Tail: Addressing Geographic Imbalance in Urban Visual Place Recognition

DGX agent

arXiv:2607.00090v1 Announce Type: cross Abstract: Urban-scale Visual Place Recognition (VPR) aims to identify the geographic location of a query image by matching it against a geo-tagged database. Whi

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

LV-ROVER: Multi-Stream Tesseract Voting for Maltese Paragraph OCR

DGX agent

arXiv:2607.00250v1 Announce Type: new Abstract: Maltese has decent text corpora and pretrained language models, but, like many languages outside the handful with large OCR benchmarks, only a single kn

model-releasesarxiv-cs-cl
2 Jul 2026
Research

Message Passing Enables Efficient Reasoning

DGX agent

arXiv:2607.01077v1 Announce Type: new Abstract: While inference-time scaling has improved the reasoning abilities of large language models (LLMs), the need to generate long chains-of-thought (CoTs) is

researcharxiv-cs-cl
2 Jul 2026
Model Releases

MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation

DGX agent

arXiv:2602.21397v2 Announce Type: replace Abstract: Prompt learning has become a dominant paradigm for adapting vision-language models (VLMs) such as CLIP to downstream tasks without modifying pretrai

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Multi-Hypothesis Test-Time Adaptation to Mitigate Underspecification

DGX agent

arXiv:2607.00259v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) seeks to improve model robustness under distribution shifts by adapting parameters using unlabeled target data. However, in

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning

DGX agent

arXiv:2607.00461v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are often constrained by a language-space bottleneck, forcing complex visual reasoning into discrete tokens whi

model-releasesarxiv-cs-cv
2 Jul 2026
Agents

NeuroFilter: Activation-Based Guardrails for Privacy-Conscious LLM Agents

DGX agent

arXiv:2601.14660v2 Announce Type: replace-cross Abstract: Agentic Large Language Models (LLMs) are models able to reason, plan, and execute tools over unstructured data. These abilities are enabling t

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

Next-Frame Decoding for Ultra-Low-Bitrate Image Compression with Video Diffusion Priors

DGX agent

arXiv:2603.15129v3 Announce Type: replace Abstract: We present a novel paradigm for ultra-low-bitrate image compression (ULB-IC) that exploits the ``temporal'' evolution in generative image compressio

model-releasesarxiv-cs-cv
2 Jul 2026
Industry

Not your weights, not your brain!

DGX agent

This post likely discusses issues of model ownership and intellectual property rights in AI development, particularly regarding concerns about who controls trained model weights and the underlying neu

industryclem-delangue--x
2 Jul 2026
Model Releases

Pano2World: End-to-End 3D Generation via Unified Multi-View Sequences

DGX agent

arXiv:2607.00832v1 Announce Type: cross Abstract: A single panorama captures the full visual sphere from one camera center, yet confines users to looking around in place without enabling true scene ex

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Prompting GPT-5 on Scrum Certification Questions: An Empirical Accuracy Study

DGX agent

arXiv:2607.00049v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in Agile Software Development for documentation, coaching, and training. As practitioners adopt the

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Retrieved Images as Visual Thought: Training-Free Multimodal In-Context Learning for the Open-vs-Closed Gap

DGX agent

arXiv:2607.00606v1 Announce Type: new Abstract: Recent work on Thinking with Images makes vision a dynamic part of reasoning, but does so through generation: the model invokes external tools, synthesi

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Right in the Right Way: LM Training with Verifiable Rewards and Human Demonstrations

DGX agent

arXiv:2607.01181v1 Announce Type: cross Abstract: RL with verifiable rewards (RLVR) has emerged as a powerful paradigm for training LMs on tasks with well-defined success metrics, such as code generat

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios

DGX agent

arXiv:2511.18011v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated powerful capabilities in general spatial understanding and reasoning. However, their fine

model-releasesarxiv-cs-cv
2 Jul 2026
Applications

Shapley in Context: Explaining Financial Language with Domain Expertise

DGX agent

arXiv:2607.00856v1 Announce Type: cross Abstract: In recent years, large language models have achieved remarkable success and have seen growing adoption in financial applications. At the same time, ex

applicationsarxiv-cs-lg
2 Jul 2026
Model Releases

Soft Mixture-of-Recursions: Going Deeper with Recursive Vision Transformers

DGX agent

arXiv:2607.00774v1 Announce Type: new Abstract: Recent recursive Transformer studies have primarily reused shared parameters across computation steps to construct compact, parameter-efficient models.

model-releasesarxiv-cs-cv
2 Jul 2026
Research

SONIC: Spectral Optimization of Noise for Inpainting with Consistency

DGX agent

arXiv:2511.19985v3 Announce Type: replace Abstract: We propose a novel training-free method for inpainting with off-the-shelf text-to-image models. While guidance-based methods in theory allow generic

researcharxiv-cs-cv
2 Jul 2026
Model Releases

SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests

DGX agent

arXiv:2607.00990v1 Announce Type: cross Abstract: Large language model (LLM)-based software engineering agents are increasingly developed to resolve software issues by generating patches from issue re

model-releasesarxiv-cs-ai
2 Jul 2026
Applications

TiRex-2: Generalizing TiRex to Multivariate Data and Streaming

DGX agent

arXiv:2607.01204v1 Announce Type: new Abstract: We introduce TiRex-2, a recurrent xLSTM-based time series foundation model that generalizes the univariate TiRex to multivariate forecasting with both p

applicationsarxiv-cs-lg
2 Jul 2026
Model Releases

Towards Metric-Agnostic Trajectory Forecasting

DGX agent

arXiv:2607.01133v1 Announce Type: new Abstract: Accurate trajectory forecasting of surrounding traffic participants is a core capability for autonomous driving, enabling vehicles to anticipate behavio

model-releasesarxiv-cs-cv
2 Jul 2026
Safety

Two AI Metrics Diverged: Will it Make All the Difference?

DGX agent

arXiv:2607.00913v1 Announce Type: new Abstract: As exponential compute scaling continues, will the capabilities of frontier AI models outstrip what is accessible to developers on a small fixed budget?

safetyarxiv-cs-ai
2 Jul 2026
Model Releases

Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations

DGX agent

arXiv:2503.13445v3 Announce Type: replace-cross Abstract: When asked to explain their decisions, LLMs can often give explanations which sound plausible to humans. But are these explanations faithful,

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

AxDafny: Agentic Verified Code Generation in Dafny

DGX agent

arXiv:2606.32007v1 Announce Type: new Abstract: We study agentic code generation in Dafny, where a model must generate both executable code and the proof artifacts for verification. We present AxDafny

model-releasesarxiv-cs-ai
1 Jul 2026
Research

Belief Contraction in Dynamic Epistemic Logic

DGX agent

arXiv:2606.31861v1 Announce Type: cross Abstract: Dynamic epistemic logic represents belief change via model transformations induced by epistemic events. Its standard formulation (Baltag, Moss, Soleck

researcharxiv-cs-ai
1 Jul 2026
Model Releases

Beyond Compilation: Evaluating Faithful Natural-Language-to-Lean Statement Formalization

DGX agent

arXiv:2606.31002v1 Announce Type: new Abstract: Theorem-proving benchmarks evaluate proof search against fixed formal statements, but natural-language-to-Lean formalization must generate the formal st

model-releasesarxiv-cs-ai
1 Jul 2026
← Previous
1…511512513514515…1369
Next →