AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,587 results
Research

Introduction to Transformers: an NLP Perspective

DGX agent

arXiv:2311.17633v2 Announce Type: replace-cross Abstract: Transformers have dominated empirical machine learning models of natural language processing. In this paper, we introduce basic concepts of Tr

researcharxiv-cs-ai
3 Jul 2026
Model Releases

LACUNA: A Testbed for Evaluating Localization Precision for LLM Unlearning

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.02513v1 Announce Type: cross Abstract: LLMs memorize sensitive training data, including personally identifiable information (PII), creating a pressing need for reliable post hoc removal met

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MMIR-TCM: Memory-Integrated Multimodal Inference and Retrieval for TCM Clinical Decision Support

DGX agent

arXiv:2607.01814v1 Announce Type: new Abstract: Traditional Chinese Medicine (TCM) diagnosis, particularly through tongue inspection, faces persistent challenges in subjectivity and reproducibility. T

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

One More Time: Revisiting Neural Quantum States from a Reinforcement Learning Perspective

DGX agent

arXiv:2607.02292v1 Announce Type: new Abstract: Neural quantum states (NQS) provide a flexible and scalable framework for approximating quantum many-body wavefunctions. Among NQS parameterizations, au

model-releasesarxiv-cs-lg
3 Jul 2026
Safety

Overthink-Triggered Slowdown Attacks on LVLM-Based Robotic Systems

DGX agent

arXiv:2607.01518v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have been increasingly integrated into robotic systems. However, these models may exhibit overthinking behaviors,

safetyarxiv-cs-ro
3 Jul 2026
Applications

PACE: A Neuro-Symbolic Framework for Plausible and Actionable Counterfactual Explanations

DGX agent

arXiv:2607.01306v1 Announce Type: new Abstract: Counterfactual explanations explain machine learning predictions by identifying minimal input changes that would alter a model's decision. Although many

applicationsarxiv-cs-ai
3 Jul 2026
Model Releases

Phonikud: Overcoming Phonetic Underspecification for Hebrew Text-To-Speech

DGX agent

arXiv:2506.12311v4 Announce Type: replace Abstract: Text-to-speech (TTS) for Modern Hebrew is challenged by the language's orthographic complexity, with existing solutions ignoring underspecified phon

model-releasesarxiv-cs-cl
3 Jul 2026
Model Releases

Population-Scale Segmentation of Penile Tissue in DIXON MRI using Deep Learning for Quantitative Phenotyping in Male Reproductive Health

DGX agent

arXiv:2607.02127v1 Announce Type: cross Abstract: Penile measurement is clinically relevant across male reproductive and urogenital health, including conditions such as micropenis, congenital and endo

model-releasesarxiv-cs-lg
3 Jul 2026
Model Releases

Scaling Trends for Lie Detector Oversight in Preference Learning

DGX agent

arXiv:2607.01567v1 Announce Type: new Abstract: Deceptive behavior in LLMs is costly to monitor and prevent, motivating approaches such as Scalable Oversight via Lie Detectors (SOLiD) (Cundy & Gleave,

model-releasesarxiv-cs-ai
3 Jul 2026
Local Ai

Self-explainable Operator Learning for Discovering Spatial Patterns in Functional Data

DGX agent

arXiv:2607.02203v1 Announce Type: new Abstract: Operator learning has emerged as a powerful tool for modeling complex physical systems in functional spaces. However, their neural network-based archite

local-aiarxiv-cs-lg
3 Jul 2026
Model Releases

Separating Expert Retention from Autonomous Source Inference in Raw-ECG-Replay-Free Continual ECG Deployment

DGX agent

arXiv:2607.01674v1 Announce Type: new Abstract: In multi-source ECG deployment, models may need to incorporate new data sources when earlier raw ECGs cannot be retained or replayed. Freezing a pretrai

model-releasesarxiv-cs-ai
3 Jul 2026
Research

SPARCLE: SPeaker-aware Aligned Representations via Contrastive Language Embeddings

DGX agent

arXiv:2607.01238v1 Announce Type: cross Abstract: Recent advances in speech synthesis have shifted from phoneme representations to direct grapheme modeling. While phonemes address the one-to-many mapp

researcharxiv-cs-ai
3 Jul 2026
Model Releases

Spec-AUF: Accept-Until-Fail Training under Train-Inference Misalignment for Masked Block Drafters

DGX agent

arXiv:2607.01893v1 Announce Type: new Abstract: Speculative decoding accelerates autoregressive generation by drafting a block of tokens that the target model verifies left-to-right, committing only t

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

SPLIT: Cross-Lingual Empathy and Cultural Grounding in English and Ukrainian LLM Responses

DGX agent

arXiv:2607.02049v1 Announce Type: cross Abstract: Large Language Models are increasingly deployed in emotional-support contexts and crisis-related situations. Nevertheless, their cross-lingual abiliti

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Token Geometry

DGX agent

arXiv:2607.01455v1 Announce Type: cross Abstract: Language models learn continuous programs over discrete symbols, with the embedding table and LM-head acting as the read/write interface between them.

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

UA-ChatDev: Uncertainty-Aware Multi-Agent Collaboration for Reliable Software Development

DGX agent

arXiv:2607.02186v1 Announce Type: new Abstract: Software development is a complex task that demands cooperation among agents with diverse roles. Large language models (LLMs) have enabled autonomous mu

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

A Filtered Mixture-of-Generators for Fully Synthetic Survival Training

DGX agent

arXiv:2607.00127v1 Announce Type: new Abstract: Survival analysis models time-to-event data, but in clinical settings training data are costly and scarce: events accrue over years of follow-up, cohort

safetyarxiv-cs-lg
2 Jul 2026
Model Releases

AGE: Adaptive-masking for Graph Embedding in Graph Retrieval-Augmented Generation

DGX agent

arXiv:2607.00052v1 Announce Type: cross Abstract: GraphRAG is an extension of retrieval-augmented generation (RAG) that supports large language models (LLMs) by referring to graph-structured data as e

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

An LLM-Based Framework for Intent-Driven Network Topology Design

DGX agent

arXiv:2607.00292v1 Announce Type: cross Abstract: Designing deployable and resilient network topologies from natural language requirements remains a challenging problem in network automation. This wor

model-releasesarxiv-cs-ai
2 Jul 2026
Local Ai

Beyond Perplexity: A Behavioral Evaluation Framework for Deployment-Memory Claims in LLM Test-Time Training

DGX agent

arXiv:2607.00368v1 Announce Type: new Abstract: Large language model test-time training (TTT) is often evaluated through local proxy metrics: models are updated on recent tokens, retrieved context, ta

local-aiarxiv-cs-cl
2 Jul 2026
Model Releases

Can Agents Generalize to the Open World? Unveiling the Fragility of Static Training in Tool Use

DGX agent

arXiv:2607.01084v1 Announce Type: new Abstract: While Large Language Model (LLM) agents demonstrate proficiency in static benchmarks, their deployment in real-world scenarios is hindered by the dynami

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Does Your ViT Still Need U-Net for Segmentation?

DGX agent

arXiv:2607.00223v1 Announce Type: new Abstract: Medical image segmentation is dominated by U-Net-style encoder-decoder architectures. Vision Transformers (ViTs) overcome the limited receptive field of

model-releasesarxiv-cs-cv
2 Jul 2026
Research

Flow-Map GRPO: Reinforcement Learning for Few-Step Flow-Map Generators via Anchored Stochastic Composition

DGX agent

arXiv:2607.00535v1 Announce Type: cross Abstract: Few-step flow-map generators, such as consistency models and MeanFlow, accelerate sampling by directly learning long-range transport maps between nois

researcharxiv-cs-ai
2 Jul 2026
Model Releases

FLYNN: Robust Neural Network for Robot Navigation using Fly Brain Topology

DGX agent

arXiv:2607.00025v1 Announce Type: cross Abstract: While deep learning models achieve state-of-the-art performance in complex tasks, they remain brittle when faced with new environments or sensory depr

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

GMO-E^2DIT: Grounded Multi-Operation Editing for E-Commerce Images

DGX agent

arXiv:2607.00920v1 Announce Type: new Abstract: Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

GPTKB v1.5: A Massive Knowledge Base for Exploring Factual LLM Knowledge

DGX agent

arXiv:2507.05740v2 Announce Type: replace Abstract: Language models are powerful artifacts, yet their factual knowledge is still poorly understood, and inaccessible to ad-hoc browsing and scalable sta

model-releasesarxiv-cs-cl
2 Jul 2026
Model Releases

GSRQ: Gain-Shape Residual Quantization for Sub-1-bit KV Cache

DGX agent

arXiv:2607.01065v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) with extended context windows is increasingly constrained by the linear growth of Key-Value (KV) cache me

model-releasesarxiv-cs-lg
2 Jul 2026
Model Releases

Homogenization of ell_2-Adversarial Training in High-Dimensions: Exact Dynamics under Stochastic Gradient Descent

DGX agent

arXiv:2607.00207v1 Announce Type: cross Abstract: We develop a framework for analyzing the learning dynamics of ell_2-adversarial training of single-index models on Gaussian mixtures in the high-dimen

model-releasesarxiv-cs-lg
2 Jul 2026
Research

Leveraging Multimodality for Real-Time Classification of Transients and Variables found by the Zwicky Transient Facility

DGX agent

arXiv:2607.00228v1 Announce Type: cross Abstract: Modern time-domain surveys such as the Zwicky Transient Facility (ZTF) generate hundreds of thousands of alerts each night, making real-time decisions

researcharxiv-cs-lg
2 Jul 2026
Model Releases

Lost in the Tail: Addressing Geographic Imbalance in Urban Visual Place Recognition

DGX agent

arXiv:2607.00090v1 Announce Type: cross Abstract: Urban-scale Visual Place Recognition (VPR) aims to identify the geographic location of a query image by matching it against a geo-tagged database. Whi

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

LV-ROVER: Multi-Stream Tesseract Voting for Maltese Paragraph OCR

DGX agent

arXiv:2607.00250v1 Announce Type: new Abstract: Maltese has decent text corpora and pretrained language models, but, like many languages outside the handful with large OCR benchmarks, only a single kn

model-releasesarxiv-cs-cl
2 Jul 2026
Research

Message Passing Enables Efficient Reasoning

DGX agent

arXiv:2607.01077v1 Announce Type: new Abstract: While inference-time scaling has improved the reasoning abilities of large language models (LLMs), the need to generate long chains-of-thought (CoTs) is

researcharxiv-cs-cl
2 Jul 2026
Model Releases

MMLoP: Multi-Modal Low-Rank Prompting for Efficient Vision-Language Adaptation

DGX agent

arXiv:2602.21397v2 Announce Type: replace Abstract: Prompt learning has become a dominant paradigm for adapting vision-language models (VLMs) such as CLIP to downstream tasks without modifying pretrai

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Multi-Hypothesis Test-Time Adaptation to Mitigate Underspecification

DGX agent

arXiv:2607.00259v1 Announce Type: cross Abstract: Test-Time Adaptation (TTA) seeks to improve model robustness under distribution shifts by adapting parameters using unlabeled target data. However, in

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Multimodal Continuous Reasoning via Asymmetric Mutual Variational Learning

DGX agent

arXiv:2607.00461v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) are often constrained by a language-space bottleneck, forcing complex visual reasoning into discrete tokens whi

model-releasesarxiv-cs-cv
2 Jul 2026
Agents

NeuroFilter: Activation-Based Guardrails for Privacy-Conscious LLM Agents

DGX agent

arXiv:2601.14660v2 Announce Type: replace-cross Abstract: Agentic Large Language Models (LLMs) are models able to reason, plan, and execute tools over unstructured data. These abilities are enabling t

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

Next-Frame Decoding for Ultra-Low-Bitrate Image Compression with Video Diffusion Priors

DGX agent

arXiv:2603.15129v3 Announce Type: replace Abstract: We present a novel paradigm for ultra-low-bitrate image compression (ULB-IC) that exploits the ``temporal'' evolution in generative image compressio

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Pano2World: End-to-End 3D Generation via Unified Multi-View Sequences

DGX agent

arXiv:2607.00832v1 Announce Type: cross Abstract: A single panorama captures the full visual sphere from one camera center, yet confines users to looking around in place without enabling true scene ex

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Prompting GPT-5 on Scrum Certification Questions: An Empirical Accuracy Study

DGX agent

arXiv:2607.00049v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in Agile Software Development for documentation, coaching, and training. As practitioners adopt the

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Retrieved Images as Visual Thought: Training-Free Multimodal In-Context Learning for the Open-vs-Closed Gap

DGX agent

arXiv:2607.00606v1 Announce Type: new Abstract: Recent work on Thinking with Images makes vision a dynamic part of reasoning, but does so through generation: the model invokes external tools, synthesi

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

Right in the Right Way: LM Training with Verifiable Rewards and Human Demonstrations

DGX agent

arXiv:2607.01181v1 Announce Type: cross Abstract: RL with verifiable rewards (RLVR) has emerged as a powerful paradigm for training LMs on tasks with well-defined success metrics, such as code generat

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

RoadBench: Benchmarking MLLMs on Fine-Grained Spatial Understanding and Reasoning under Urban Road Scenarios

DGX agent

arXiv:2511.18011v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have demonstrated powerful capabilities in general spatial understanding and reasoning. However, their fine

model-releasesarxiv-cs-cv
2 Jul 2026
Applications

Shapley in Context: Explaining Financial Language with Domain Expertise

DGX agent

arXiv:2607.00856v1 Announce Type: cross Abstract: In recent years, large language models have achieved remarkable success and have seen growing adoption in financial applications. At the same time, ex

applicationsarxiv-cs-lg
2 Jul 2026
Model Releases

Soft Mixture-of-Recursions: Going Deeper with Recursive Vision Transformers

DGX agent

arXiv:2607.00774v1 Announce Type: new Abstract: Recent recursive Transformer studies have primarily reused shared parameters across computation steps to construct compact, parameter-efficient models.

model-releasesarxiv-cs-cv
2 Jul 2026
Research

SONIC: Spectral Optimization of Noise for Inpainting with Consistency

DGX agent

arXiv:2511.19985v3 Announce Type: replace Abstract: We propose a novel training-free method for inpainting with off-the-shelf text-to-image models. While guidance-based methods in theory allow generic

researcharxiv-cs-cv
2 Jul 2026
Model Releases

SWE-Doctor: Guiding Software Engineering Agents with Runtime Diagnosis from Multi-Faceted Bug Reproduction Tests

DGX agent

arXiv:2607.00990v1 Announce Type: cross Abstract: Large language model (LLM)-based software engineering agents are increasingly developed to resolve software issues by generating patches from issue re

model-releasesarxiv-cs-ai
2 Jul 2026
Applications

TiRex-2: Generalizing TiRex to Multivariate Data and Streaming

DGX agent

arXiv:2607.01204v1 Announce Type: new Abstract: We introduce TiRex-2, a recurrent xLSTM-based time series foundation model that generalizes the univariate TiRex to multivariate forecasting with both p

applicationsarxiv-cs-lg
2 Jul 2026
Model Releases

Towards Metric-Agnostic Trajectory Forecasting

DGX agent

arXiv:2607.01133v1 Announce Type: new Abstract: Accurate trajectory forecasting of surrounding traffic participants is a core capability for autonomous driving, enabling vehicles to anticipate behavio

model-releasesarxiv-cs-cv
2 Jul 2026
← Previous
1…404405406407408…1075
Next →