AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlog
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,676 results
Model Releases

Information-Aware KV Cache Compression for Long Reasoning

DGX agent

arXiv:2606.26875v1 Announce Type: cross Abstract: Reasoning capability has advanced rapidly in large language models (LLMs), leading to an increasing size of key-value (KV) cache in both prefilling an

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Inherited Circuits, Learned Semantics: How Fine-Tuning Creates Evasion Vulnerabilities Invisible to Standard Evaluation

DGX agent

arXiv:2606.27091v1 Announce Type: cross Abstract: LLMs fine-tuned for security classification are usually evaluated on held-out examples from the same distribution as their training data. We show that

model-releasesarxiv-cs-ai
26 Jun 2026
Research

Learning from Equivalence Queries, Revisited

DGX agent

arXiv:2604.04535v2 Announce Type: replace Abstract: Modern machine learning systems, such as generative models and recommendation systems, often evolve through a cycle of deployment, user interaction,

researcharxiv-cs-lg
26 Jun 2026
Model Releases

Learning to Select Maximum Clique Algorithms: From Traditional Machine Learning to a Dual-Channel Hybrid Neural Architecture

DGX agent

arXiv:2508.08005v4 Announce Type: replace-cross Abstract: The Maximum Clique Problem (MCP) is an NP-hard problem with wide-ranging applications in fields such as bioinformatics, network science, and s

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Look-Before-Move: Narrative-Grounded World Visual Attention in Dynamic 3D Story Worlds

DGX agent

arXiv:2606.26964v1 Announce Type: new Abstract: As embodied AI and world models increasingly operate in dynamic 3D environments, visual perception must move beyond passively interpreting given observa

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Mask to Concept: Auto-Promptable SAM3 via Efficient Test-Time Concept Embedding Search for Few-Shot Annotation

DGX agent

arXiv:2606.26711v1 Announce Type: new Abstract: Transforming foundation segmentation models from human-prompted tools into auto-promptable annotators is critical for scalable medical data annotation.

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

MKG-RAG-Bench: Benchmarking Retrieval in Multimodal Knowledge Graph-Augmented Generation

DGX agent

arXiv:2606.26458v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) over knowledge graphs has emerged as a promising approach for grounding large language models, yet existing benchma

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

OpenAI says GPT-5.6 Sol and Terra were capable of identifying vulnerabilities but were unable to execute autonomous, end-to-end attacks against hardened targets (OpenAI)

DGX agent

OpenAI: OpenAI says GPT-5.6 Sol and Terra were capable of identifying vulnerabilities but were unable to execute autonomous, end-to-end attacks against hardened targets — GPT-5.6 is a new family of th

model-releasestechmeme
26 Jun 2026
Model Releases

OpenFinGym: A Verifiable Multi-Task Gym Environment for Evaluating Quant Agents

DGX agent

arXiv:2606.26350v1 Announce Type: new Abstract: Although large language model agents are increasingly applied to quantitative-finance workflows, their evaluation remains fragmented across isolated tas

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Parametric Open Source Games

DGX agent

arXiv:2606.27068v1 Announce Type: cross Abstract: Open-source game theory studies agents whose behavior may depend on one another's decision procedures, but most existing models use discrete or symbol

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

PersistentKV: Page-Aware Decode Scheduling for Long-Context LLM Serving on Commodity GPUs

DGX agent

arXiv:2606.26666v1 Announce Type: new Abstract: Autoregressive large language model (LLM) serving is increasingly limited by key-value (KV) cache movement rather than dense matrix multiplication. Mode

model-releasesarxiv-cs-lg
26 Jun 2026
Applications

PhysiFormer: Learning to Simulate Mechanics in World Space

DGX agent

arXiv:2606.27364v1 Announce Type: new Abstract: We present PhysiFormer, a diffusion transformer for physically-plausible 3D object motion. Unlike video world models that operate in view-dependent pixe

applicationsarxiv-cs-cv
26 Jun 2026
Model Releases

ReasonCLIP-58M: Visually Grounded Commonsense Reasoning Supervision for CLIP

DGX agent

arXiv:2606.26794v1 Announce Type: cross Abstract: CLIP and its variants are widely adopted visual backbones in multimodal systems, but their pretraining remains dominated by descriptive image-text ali

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

ReportLogic: Evaluating Logical Quality in Deep Research Reports

DGX agent

arXiv:2602.18446v2 Announce Type: replace-cross Abstract: Users increasingly rely on Large Language Models (LLMs) for Deep Research, using them to synthesize diverse sources into structured reports th

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

SamaVaani: Auditing and Debiasing Multilingual Clinical ASR for Indian Languages

DGX agent

arXiv:2606.26901v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) is increasingly used to document clinical encounters, yet its reliability in multilingual and demographically diver

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

See & Sniff: Learning Visuo-Olfactory Representations

DGX agent

arXiv:2606.27307v1 Announce Type: new Abstract: While modern multimodal models integrate vision with language, audio, or touch, olfaction remains largely unexplored due to the lack of paired visuo-olf

model-releasesarxiv-cs-cv
26 Jun 2026
Model Releases

ShareLock: A Stealthy Multi-Tool Threshold Poisoning Attack Against MCP

DGX agent

arXiv:2606.27027v1 Announce Type: cross Abstract: With the rapid evolution of LLM-driven agents, Model Context Protocol (MCP), an open protocol bridging LLMs with external tools, has quickly become fo

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

Staying VIGILant: Mitigating Visual Laziness via Counterfactual Visual Alignment in MLLMs

DGX agent

arXiv:2606.26387v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) extend large language models (LLMs) with visual perception, enabling joint reasoning over images and text. De

safetyarxiv-cs-cl
26 Jun 2026
Tutorials

Structure Before Collapse: Transient semantic geometry in next-token prediction

DGX agent

arXiv:2606.26749v1 Announce Type: cross Abstract: Neural Collapse predicts that balanced one-hot classification pushes model representations to be equally far from each other; a symmetric configuratio

tutorialsarxiv-cs-cl
26 Jun 2026
Model Releases

Towards Video Anomaly Detection from Event Streams: A Baseline and Benchmark Datasets

DGX agent

arXiv:2603.24991v2 Announce Type: replace Abstract: Event-based vision, characterized by low redundancy, focus on dynamic motion, and inherent privacy-preserving properties, naturally fits the demands

model-releasesarxiv-cs-cv
26 Jun 2026
Research

When are likely answers right? On Sequence Probability and Correctness in LLMs

DGX agent

arXiv:2606.27359v1 Announce Type: cross Abstract: Many decoding methods for large language models can be understood as shifting probability mass toward outputs that are more likely under the model, ei

researcharxiv-cs-lg
26 Jun 2026
Model Releases

Which one is Claude is pretty obvious. GLM-5.2 is a beast in some ways, but doesn't have the self-reflective persona of Claude, and isn't re…

DGX agent

Ethan Mollick compares Claude and GLM-5.2 AI models, noting that while GLM-5.2 excels in certain capabilities, Claude distinguishes itself through its self-reflective persona and other characteristics

model-releasesethan-mollick--x
26 Jun 2026
Model Releases

Zero-Shot Size Transfer for Neural ODEs on Sparse Random Graphs: Graphon Limits and Adjoint Convergence

DGX agent

arXiv:2606.26662v1 Announce Type: cross Abstract: Graph Neural Differential Equations (GNDEs) model continuous-time graph dynamics by parameterizing Neural ODE velocity fields with Graph Neural Networ

model-releasesarxiv-cs-ai
26 Jun 2026
Research

Auto-Configured Explainable Graph Neural Networks for Multi-Site Pollution Prediction

DGX agent

arXiv:2606.24978v1 Announce Type: new Abstract: Accurate particulate matter (PM) prediction is crucial for mitigating air pollution. Graph Neural Networks (GNNs) effectively model spatiotemporal depen

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Beyond Function Calling: Benchmarking Tool-Using Agents under Tool-Environment Unreliability

DGX agent

arXiv:2606.25819v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that solve tasks by interacting with external tool environments. Although recent tool-use benc

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Blasto-Net: An Explainable Multi-Task Learning for Blastocyst Segmentation, Grading, and Implantation Prediction

DGX agent

arXiv:2606.25463v1 Announce Type: cross Abstract: This study introduces Blasto-Net, a multi-task deep learning model for comprehensive blastocyst analysis. The proposed model performs three tasks simu

researcharxiv-cs-lg
25 Jun 2026
Model Releases

Constraint Tax in Open-Weight LLMs: An Empirical Study of Tool Calling Suppression Under Structured Output Constraints

DGX agent

arXiv:2606.25605v1 Announce Type: new Abstract: Tool Calling and Structured Output are two core capabilities of modern Agent systems, yet their interaction under joint deployment conditions remains in

model-releasesarxiv-cs-cl
25 Jun 2026
Research

DFMU: Data-Frugal Machine Unlearning

DGX agent

arXiv:2606.25410v1 Announce Type: new Abstract: Machine unlearning is an emerging domain that ensures the safe removal of elements (includes concepts, attributes, entity and class) from the trained mo

researcharxiv-cs-lg
25 Jun 2026
Research

Emergent Capabilities Arise Randomly from Learning Sparse Attention Patterns

DGX agent

arXiv:2606.25010v1 Announce Type: cross Abstract: Neural scaling laws for transformer language models predict smooth improvements in pretraining loss with increasing parameters, but downstream capabil

researcharxiv-cs-cl
25 Jun 2026
Model Releases

Enhancing Pathological VLMs with Cross-scale Reasoning

DGX agent

arXiv:2606.17412v3 Announce Type: replace Abstract: Pathological images are inherently multi-scale, requiring pathologists to integrate evidence from global tissue architecture at low magnification to

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

Evaluating LLMs on Real-World Software Performance Optimization

DGX agent

arXiv:2606.25530v1 Announce Type: cross Abstract: Software performance optimization is a notoriously complex and manual task. Despite the growing use of Large Language Models (LLMs) for code refinemen

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

HG-Bench: A Benchmark for Multi-Page Handwritten Answer-Region Grounding in Automated Homework Assessment

DGX agent

arXiv:2606.25491v1 Announce Type: new Abstract: Automated homework assessment depends not only on recognizing student answers, but also on accurately locating where each answer and each intermediate r

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

Homogeneity Bias in Open-Weight LLMs Is Robust to Decoding Hyperparameters

DGX agent

arXiv:2501.02211v2 Announce Type: replace-cross Abstract: Large language models (LLMs) reproduce homogeneity bias -- the tendency to portray marginalized groups as more internally similar than dominan

safetyarxiv-cs-cl
25 Jun 2026
Model Releases

Invoice Haystack: Benchmarking Document Retrieval and Visual Question Answering Under Strong Visual Homogeneity

DGX agent

arXiv:2606.25343v1 Announce Type: new Abstract: Vision Language Models have achieved near-human performance on single-document Visual Question Answering, yet their effectiveness degrades significantly

model-releasesarxiv-cs-cv
25 Jun 2026
Hardware

JetSpec: Breaking the Scaling Ceiling of Speculative Decoding with Parallel Tree Drafting

DGX agent

arXiv:2606.18394v2 Announce Type: replace Abstract: Speculative decoding (SD) accelerates autoregressive Large Language Models (LLMs) by drafting multiple tokens and verifying them in parallel, but it

hardwarearxiv-cs-cl
25 Jun 2026
Applications

LLM Evolution as an Industry-Scale Ecosystem: A Lifecycle Perspective on Continual Learning

DGX agent

arXiv:2606.24901v1 Announce Type: new Abstract: Continual learning capability is critical for Industrial LLMs, as deployed models must be continuously updated to meet evolving requirements and environ

applicationsarxiv-cs-lg
25 Jun 2026
Model Releases

MacroLens: A Multi-Task Benchmark for Contextual Financial Reasoning under Macroeconomic Scenarios

DGX agent

arXiv:2606.24950v1 Announce Type: new Abstract: Financial decision-making is contextual: forecasting prices, valuing companies, and assessing event exposure weigh price history, accounting fundamental

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Multilingual Hematology Visual Question Answering Dataset

DGX agent

arXiv:2606.25246v1 Announce Type: cross Abstract: Vision Language Models (VLMs) have shown promising capabilities in medical image analysis by jointly understanding visual and textual information for

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

OpenAI will delay GPT-5.6 after Trump administration request

DGX agent

The Trump administration, apprehensive of potential security issues, has reportedly asked OpenAI to stagger the release of its next big-ticket model, GPT-5.6. The Information reported that OpenAI CEO

model-releasesthe-verge-ai
25 Jun 2026
Research

@OpenRouter https://openrouter.ai/sakana/fugu-ultra

DGX agent

OpenRouter is a platform that provides API access to multiple large language models and AI services through a unified interface. Sakana AI's Fugu-Ultra appears to be an AI model available through Open

researchdavid-ha--x
25 Jun 2026
Model Releases

Pulmonary Embolism Risk Stratification from CTPA and Medical Records: Vascular Graphs Are Not All You Need

DGX agent

arXiv:2606.25956v1 Announce Type: new Abstract: Risk stratification for pulmonary embolism (PE) is critical for clinical decision-making. Stratification guidelines are based on patient medical records

model-releasesarxiv-cs-cv
25 Jun 2026
Applications

Quantifying Explainable AI-introduced signal noise on ECG data with Spectral Entropy

DGX agent

arXiv:2606.24974v1 Announce Type: new Abstract: Explainability techniques are used to assess the output of various deep learning models. This is especially true in healthcare, where models need to be

applicationsarxiv-cs-lg
25 Jun 2026
Model Releases

RoboAtlas: Contextual Active SLAM

DGX agent

arXiv:2606.26046v1 Announce Type: cross Abstract: We present RoboAtlas, a contextual Active SLAM framework that adaptively balances geometric exploration and semantic reasoning using a scalable 3D sem

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

RoboRouter: Training-Free Policy Routing for Robotic Manipulation

DGX agent

arXiv:2603.07892v4 Announce Type: replace Abstract: Research on robotic manipulation has developed a diverse set of policy paradigms, including vision-language-action (VLA) models, vision-action (VA)

model-releasesarxiv-cs-ro
25 Jun 2026
Model Releases

Sarashina2.2-TTS: Tackling Kanji Polyphony in Japanese Speech Generation via Data Scaling and Targeted Data Synthesis

DGX agent

arXiv:2606.25369v1 Announce Type: cross Abstract: While large language model (LLM)-based text-to-speech (TTS) systems have achieved high-quality speech synthesis, most existing systems focus on Englis

model-releasesarxiv-cs-cl
25 Jun 2026
Research

Scale or Reason? A Compute-Equivalent Analysis of Reasoning Distillation

DGX agent

arXiv:2509.22193v2 Announce Type: replace Abstract: Distilling reasoning traces from strong teacher models has become the standard recipe for building capable small language models. Yet reasoning trac

researcharxiv-cs-cl
25 Jun 2026
Model Releases

SFL-MTSC: Leveraging Semantic Frame-Level Multi-Task Self-Consistency for Robust Multi-Intent Spoken Language Understanding

DGX agent

arXiv:2606.25552v1 Announce Type: new Abstract: Prompt-based spoken language understanding (SLU) with large language models (LLMs) often suffers from inconsistent intent--slot structures due to decodi

model-releasesarxiv-cs-cl
25 Jun 2026
Model Releases

ShutterMuse: Capture-Time Photography Guidance with MLLMs

DGX agent

arXiv:2606.25763v1 Announce Type: new Abstract: Real-world photography requires capture-time guidance for both camera framing and subject pose. Yet existing aesthetic cropping benchmarks mainly evalua

model-releasesarxiv-cs-cv
25 Jun 2026
← Previous
1…515516517518519…1369
Next →