AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Model Releases

Trusted access for the next era of cyber defense

DGX agent

OpenAI is expanding trusted access to its AI models for cybersecurity purposes, enabling vetted researchers, defenders, and organizations to leverage advanced AI capabilities for cyber defense applica

model-releasesopenai
14 Apr 2026
Safety

Weird Generalization is Weirdly Brittle

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2604.10022v1 Announce Type: new Abstract: Weird generalization is a phenomenon in which models fine-tuned on data from a narrow domain (e.g. insecure code) develop surprising traits that manifes

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

What and Where to Adapt: Structure-Semantics Co-Tuning for Machine Vision Compression via Synergistic Adapters

DGX agent

arXiv:2604.10017v1 Announce Type: new Abstract: Parameter-efficient fine-tuning of pre-trained codecs is a promising direction in image compression for human and machine vision. While most existing wo

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

What Factors Affect LLMs and RLLMs in Financial Question Answering?

DGX agent

arXiv:2507.08339v4 Announce Type: replace Abstract: Recently, large language models (LLMs) and reasoning large language models (RLLMs) have gained considerable attention from many researchers. RLLMs e

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

DGX agent

arXiv:2604.10101v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creati

model-releasesarxiv-cs-cl
14 Apr 2026
Research

XD-MAP: Cross-Modal Domain Adaptation via Semantic Parametric Maps for Scalable Training Data Generation

DGX agent

arXiv:2601.14477v2 Announce Type: replace-cross Abstract: Until open-world foundation models match the performance of specialized approaches, deep learning systems remain dependent on task- and sensor

researcharxiv-cs-ai
14 Apr 2026
Model Releases

4D-RGPT: Toward Region-level 4D Understanding via Perceptual Distillation

DGX agent

arXiv:2512.17012v3 Announce Type: replace Abstract: Despite advances in Multimodal LLMs (MLLMs), their ability to reason over 3D structures and temporal dynamics remains limited, constrained by weak 4

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Across the Levels of Analysis: Explaining Predictive Processing in Humans Requires More Than Machine-Estimated Probabilities

DGX agent

arXiv:2604.09466v1 Announce Type: new Abstract: Under the lens of Marr's levels of analysis, we critique and extend two claims about language models (LMs) and language processing: first, that predicti

researcharxiv-cs-cl
13 Apr 2026
Model Releases

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments

DGX agent

arXiv:2604.06111v2 Announce Type: replace Abstract: Existing Agent benchmarks suffer from two critical limitations: high environment interaction overhead (up to 41% of total evaluation time) and imbal

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Automated Instruction Revision (AIR): A Structured Comparison of Task Adaptation Strategies for LLM

DGX agent

arXiv:2604.09418v1 Announce Type: new Abstract: This paper studies Automated Instruction Revision (AIR), a rule-induction-based method for adapting large language models (LLMs) to downstream tasks usi

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine

DGX agent

arXiv:2603.06665v2 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) often benefit from chain-of-thought (CoT) prompting in general domains, yet its efficacy in medical vision

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding

DGX agent

arXiv:2511.23071v2 Announce Type: replace-cross Abstract: Reading scene text, that is, text appearing in images, has numerous application areas, including assistive technology, search, and e-commerce.

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Boosting Brain-inspired Path Integration Efficiency via Learning-based Replication of Continuous Attractor Neurodynamics

DGX agent

arXiv:2511.17687v2 Announce Type: replace Abstract: The brain's Path Integration (PI) mechanism offers substantial guidance and inspiration for Brain-Inspired Navigation (BIN). However, the PI capabil

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

CAD 100K: A Comprehensive Multi-Task Dataset for Car Related Visual Anomaly Detection

DGX agent

arXiv:2604.09023v1 Announce Type: new Abstract: Multi-task visual anomaly detection is critical for car-related manufacturing quality assessment. However, existing methods remain task-specific, hinder

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

DeFakeQ: Enabling Real-Time Deepfake Detection on Edge Devices via Adaptive Bidirectional Quantization

DGX agent

arXiv:2604.08847v1 Announce Type: new Abstract: Deepfake detection has become a fundamental component of modern media forensics. Despite significant progress in detection accuracy, most existing metho

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs

DGX agent

arXiv:2604.08846v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have been shown to be vulnerable to malicious queries that can elicit unsafe responses. Recent work uses prom

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Fine-Grained Action Segmentation for Renorrhaphy in Robot-Assisted Partial Nephrectomy

DGX agent

arXiv:2604.09051v1 Announce Type: new Abstract: Fine-grained action segmentation during renorrhaphy in robot-assisted partial nephrectomy requires frame-level recognition of visually similar suturing

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

FIRE-CIR: Fine-grained Reasoning for Composed Fashion Image Retrieval

DGX agent

arXiv:2604.09114v1 Announce Type: new Abstract: Composed image retrieval (CIR) aims to retrieve a target image that depicts a reference image modified by a textual description. While recent vision-lan

model-releasesarxiv-cs-cv
13 Apr 2026
Applications

Folks, this is not Jevon's Paradox. this is just normal supply and demand. It turns out the utility of AI is high enough that people have hi…

DGX agent

Folks, this is not Jevon's Paradox. this is just normal supply and demand. It turns out the utility of AI is high enough that people have high demand, which is outstripping supply (so prices will go u

applicationsethan-mollick--x
13 Apr 2026
Safety

GRM: Utility-Aware Jailbreak Attacks on Audio LLMs via Gradient-Ratio Masking

DGX agent

arXiv:2604.09222v1 Announce Type: cross Abstract: Audio large language models (ALLMs) enable rich speech-text interaction, but they also introduce jailbreak vulnerabilities in the audio modality. Exis

safetyarxiv-cs-ai
13 Apr 2026
Applications

Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench…

DGX agent

Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench Reasoning • #1 on BridgeBench Speed • #1 on BridgeBench Low

applicationselon-musk--x
13 Apr 2026
Research

How Noise Benefits AI-generated Image Detection

DGX agent

arXiv:2511.16136v2 Announce Type: replace Abstract: The rapid advancement of generative models has made real and synthetic images increasingly indistinguishable. Although extensive efforts have been d

researcharxiv-cs-cv
13 Apr 2026
Hardware

Is an nvidia DGK Spark or similar worth it?

DGX agent

This Reddit thread on r/ollama discusses whether the NVIDIA DGX Spark — powered by the GB10 Grace Blackwell Superchip and delivering 1 petaFLOP of performance — is a worthwhile investment for running

hardwarer-ollama
13 Apr 2026
Research

Is More Data Worth the Cost? Dataset Scaling Laws in a Tiny Attention-Only Decoder

DGX agent

arXiv:2604.09389v1 Announce Type: cross Abstract: Training Transformer language models is expensive, as performance typically improves with increasing dataset size and computational budget. Although s

researcharxiv-cs-cl
13 Apr 2026
Applications

LLM-Rosetta: A Hub-and-Spoke Intermediate Representation for Cross-Provider LLM API Translation

DGX agent

arXiv:2604.09360v1 Announce Type: cross Abstract: The rapid proliferation of Large Language Model (LLM) providers--each exposing proprietary API formats--has created a fragmented ecosystem where appli

applicationsarxiv-cs-ai
13 Apr 2026
Model Releases

Looking for people with different hardware to help benchmark local LLM behavioral reliability

DGX agent

A Reddit post in the r/ollama community seeking volunteers with diverse hardware setups to participate in a collaborative effort to benchmark the **behavioral reliability** of locally-run large langua

model-releasesr-ollama
13 Apr 2026
Model Releases

MARINER: A 3E-Driven Benchmark for Fine-Grained Perception and Complex Reasoning in Open-Water Environments

DGX agent

arXiv:2604.08615v1 Announce Type: cross Abstract: Fine-grained visual understanding and high-level reasoning in real-world open-water environments remain under-explored due to the lack of dedicated be

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Measurement-Consistent Langevin Corrector for Stabilizing Latent Diffusion Inverse Problem Solvers

DGX agent

arXiv:2601.04791v3 Announce Type: replace Abstract: While latent diffusion models (LDMs) have emerged as powerful priors for inverse problems, existing LDM-based solvers frequently suffer from instabi

researcharxiv-cs-cv
13 Apr 2026
Safety

Mechanisms of Introspective Awareness

DGX agent

arXiv:2603.21396v2 Announce Type: replace Abstract: Recent work has shown that LLMs can sometimes detect when steering vectors are injected into their residual stream and identify the injected concept

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭…

DGX agent

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭𝐡𝐞 𝐩𝐫𝐨𝐩𝐞𝐫 𝐢𝐧𝐭𝐞𝐫𝐚𝐜𝐭𝐢𝐨𝐧 𝐛𝐞𝐭𝐰𝐞𝐞𝐧 𝐡𝐚𝐫𝐧𝐞𝐬𝐬 𝐚𝐧𝐝 𝐦𝐞𝐦𝐨𝐫𝐲, 𝐦𝐞𝐦𝐨𝐫𝐲 𝐚𝐥𝐨𝐧𝐞 𝐢𝐬 𝐩𝐨

model-releasesharrison-chase--x
13 Apr 2026
Model Releases

MolPaQ: Modular Quantum-Classical Patch Learning for Interpretable Molecular Generation

DGX agent

arXiv:2604.08575v1 Announce Type: cross Abstract: Molecular generative models must jointly ensure validity, diversity, and property control, yet existing approaches typically trade off among these obj

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Mosaic: Multimodal Jailbreak against Closed-Source VLMs via Multi-View Ensemble Optimization

DGX agent

arXiv:2604.09253v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are powerful but remain vulnerable to multimodal jailbreak attacks. Existing attacks mainly rely on either explicit visu

safetyarxiv-cs-ai
13 Apr 2026
Agents

MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation

DGX agent

arXiv:2604.08782v1 Announce Type: new Abstract: Large language models (LLMs) suffer significant performance degradation when user instructions and context are distributed over multiple conversational

agentsarxiv-cs-cl
13 Apr 2026
Model Releases

Multivariate Time Series Anomaly Detection via Dual-Branch Reconstruction and Autoregressive Flow-based Residual Density Estimation

DGX agent

arXiv:2604.08582v1 Announce Type: cross Abstract: Multivariate Time Series Anomaly Detection (MTSAD) is critical for real-world monitoring scenarios such as industrial control and aerospace systems. M

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Natural Riemannian gradient for learning functional tensor networks

DGX agent

arXiv:2604.09263v1 Announce Type: cross Abstract: We consider machine learning tasks with low-rank functional tree tensor networks (TTN) as the learning model. While in the case of least-squares regre

model-releasesarxiv-cs-lg
13 Apr 2026
Safety

Neural Distribution Prior for LiDAR Out-of-Distribution Detection

DGX agent

arXiv:2604.09232v1 Announce Type: cross Abstract: LiDAR-based perception is critical for autonomous driving due to its robustness to poor lighting and visibility conditions. Yet, current models operat

safetyarxiv-cs-ai
13 Apr 2026
Local Ai

Neurons Speak in Ranges: Breaking Free from Discrete Neuronal Attribution

DGX agent

arXiv:2502.06809v3 Announce Type: replace-cross Abstract: Pervasive polysemanticity in large language models (LLMs) undermines discrete neuron-concept attribution, posing a significant challenge for m

local-aiarxiv-cs-ai
13 Apr 2026
Local Ai

Offline-First LLM Architecture for Adaptive Learning in Low-Connectivity Environments

DGX agent

arXiv:2603.03339v5 Announce Type: replace-cross Abstract: Artificial intelligence (AI) and large language models (LLMs) are transforming educational technology by enabling conversational tutoring, per

local-aiarxiv-cs-cl
13 Apr 2026
Model Releases

Ollama 0.20.6 is here with improved Gemma 4 tool calling! more improvements to come for Gemma 4!

DGX agent

Ollama version 0.20.6 has been released, featuring improved tool calling support for Google's Gemma 4 model. The update focuses on enhancing the reliability and functionality of function/tool calling

model-releasesollama--x
13 Apr 2026
Research

p1: Better Prompt Optimization with Fewer Prompts

DGX agent

arXiv:2604.08801v1 Announce Type: cross Abstract: Prompt optimization improves language models without updating their weights by searching for a better system prompt, but its effectiveness varies wide

researcharxiv-cs-cl
13 Apr 2026
Model Releases

Pretrain-then-Adapt: Uncertainty-Aware Test-Time Adaptation for Text-based Person Search

DGX agent

arXiv:2604.08598v1 Announce Type: cross Abstract: Text-based person search faces inherent limitations due to data scarcity, driven by stringent privacy constraints and the high cost of manual annotati

model-releasesarxiv-cs-cv
13 Apr 2026
Safety

Rethinking Prospect Theory for LLMs: Revealing the Instability of Decision-Making under Epistemic Uncertainty

DGX agent

arXiv:2508.08992v3 Announce Type: replace Abstract: Prospect Theory (PT) models human decision-making behaviour under uncertainty, among which linguistic uncertainty is commonly adopted in real-world

safetyarxiv-cs-ai
13 Apr 2026
Safety

SafeMind: A Risk-Aware Differentiable Control Framework for Adaptive and Safe Quadruped Locomotion

DGX agent

arXiv:2604.09474v1 Announce Type: cross Abstract: Learning-based quadruped controllers achieve impressive agility but typically lack formal safety guarantees under model uncertainty, perception noise,

safetyarxiv-cs-ai
13 Apr 2026
Safety

Scene-Agnostic Object-Centric Representation Learning for 3D Gaussian Splatting

DGX agent

arXiv:2604.09045v1 Announce Type: new Abstract: Recent works on 3D scene understanding leverage 2D masks from visual foundation models (VFMs) to supervise radiance fields, enabling instance-level 3D s

safetyarxiv-cs-cv
13 Apr 2026
Model Releases

Seeing is Believing: Robust Vision-Guided Cross-Modal Prompt Learning under Label Noise

DGX agent

arXiv:2604.09532v1 Announce Type: cross Abstract: Prompt learning is a parameter-efficient approach for vision-language models, yet its robustness under label noise is less investigated. Visual conten

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Spectral Geometry of LoRA Adapters Encodes Training Objective and Predicts Harmful Compliance

DGX agent

arXiv:2604.08844v1 Announce Type: new Abstract: We study whether low-rank spectral summaries of LoRA weight deltas can identify which fine-tuning objective was applied to a language model, and whether

model-releasesarxiv-cs-lg
13 Apr 2026
Research

Streaming Video Instruction Tuning

DGX agent

arXiv:2512.21334v2 Announce Type: replace Abstract: We present Streamo, a real-time streaming video LLM that serves as a general-purpose interactive assistant. Unlike existing online video models that

researcharxiv-cs-cv
13 Apr 2026
Model Releases

TaxPraBen: A Scalable Benchmark for Structured Evaluation of LLMs in Chinese Real-World Tax Practice

DGX agent

arXiv:2604.08948v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel in various general domains, they exhibit notable gaps in the highly specialized, knowledge-intensive, and legal

model-releasesarxiv-cs-cl
13 Apr 2026
← Previous
1…561562563564565…1371
Next →