AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlog
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,597 results
Model Releases

When Synthetic Users Fail: A Cross-Domain Benchmark of LLM-Simulated Human Survey Responses

DGX agent

arXiv:2607.26348v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as synthetic users, stand-ins for human respondents whose simulated answers feed product, policy, and

model-releasesarxiv-cs-cl
30 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tutorials

A Causality-aware Infer-diagnose-refine Framework for Test-time Modality Adaptation in VLA Models

DGX agent

arXiv:2607.25516v1 Announce Type: new Abstract: Vision-language-action (VLA) models predict sequential actions to execute tasks specified by language instructions, conditioned on visual observations a

tutorialsarxiv-cs-ro
29 Jul 2026
Research

A Path Integral Model of Cognition

DGX agent

arXiv:2607.24807v1 Announce Type: cross Abstract: We develop the mathematical and physical formulation of cognitive cost optimization that underlies the path-integral model of consciousness. The goal-

researcharxiv-cs-ai
29 Jul 2026
Research

CADENCE: A Cardiac Atom Dictionary for Interpretable Neural Concept Extraction from ECG Foundation Models

DGX agent

arXiv:2607.25244v1 Announce Type: new Abstract: Foundation models for 12-lead electrocardiograms (ECGs) transfer well across clinical tasks, but the physiological knowledge encoded in their representa

researcharxiv-cs-ai
29 Jul 2026
Tutorials

Diffusion Disambiguation Models for Partial Label Learning

DGX agent

arXiv:2507.00411v2 Announce Type: replace Abstract: Learning from ambiguous labels is a long-standing problem in practical machine learning applications. The purpose of partial label learning (PLL) is

tutorialsarxiv-cs-lg
29 Jul 2026
Research

DS@GT ARC at CheckThat! 2026: LLM-Based Trace Ranking and Grouped Reward Modeling for Multilingual Numerical Claim Verification

DGX agent

arXiv:2607.25069v1 Announce Type: cross Abstract: Automated verification of numerical claims is a challenging problem, as it requires both language understanding and quantitative reasoning. This paper

researcharxiv-cs-ai
29 Jul 2026
Research

FLASH: Efficient Impact Fall Detection with Unified Hypergraph State-Space Model

DGX agent

arXiv:2607.25791v1 Announce Type: new Abstract: Falls represent a critical public health challenge, and accurate detection of the impact moment when an individual hits the ground is crucial for timely

researcharxiv-cs-cv
29 Jul 2026
Research

Foundation Models for EEG Are Blind to Long-Range Temporal Correlations: A Spectral-Temporal Dissociation Behind Their Cross-Population Fragility

DGX agent

arXiv:2607.24834v1 Announce Type: cross Abstract: Objective. Electroencephalography (EEG) foundation models (FMs) are trained to reconstruct or contrastively align short patches, then pooled into a fi

researcharxiv-cs-ai
29 Jul 2026
Model Releases

I tried running a 1.56TB MoE model on a 6GB RTX 4050 Laptop, Here’s the result

DGX agent

The Test Bench Setup I tested running a massive 1.56TB Mixture-of-Experts (MoE) checkpoint (96 shards, 93 layers, 896 experts/layer, ~4.46 bits/param MXFP4) on a budget gaming laptop. Laptop: HP Victu

model-releasesr-localllama
29 Jul 2026
Local Ai

Linear-LLM-SCM: Benchmarking LLMs for Coefficient Elicitation in Linear-Gaussian Causal Models

DGX agent

arXiv:2602.10282v2 Announce Type: replace Abstract: Large language models (LLMs) have shown potential in identifying qualitative causal relations, but their ability to perform quantitative causal reas

local-aiarxiv-cs-lg
29 Jul 2026
Applications

MEDIC-AD: Towards Medical Vision-Language Model's Clinical Intelligence

DGX agent

arXiv:2603.27176v2 Announce Type: replace Abstract: Lesion detection, symptom tracking, and visual explainability are central to real-world medical image analysis, yet current medical Vision-Language

applicationsarxiv-cs-cv
29 Jul 2026
Model Releases

Neurai-VN Benchmark: Standardized Machine Learning Models for Multimodal Digital Phenotyping in Mental Health Classification

DGX agent

arXiv:2607.25232v1 Announce Type: new Abstract: Digital phenotyping (DP) using smartphones and wearable devices has shown considerable potential for mental health monitoring. However, progress remains

model-releasesarxiv-cs-lg
29 Jul 2026
Safety

NEXT: Reasoning-Driven Video Recommendation via a Vision-Language Model

DGX agent

arXiv:2607.24789v1 Announce Type: cross Abstract: We present NEXT (Next-interest EXploration Transformer), a reasoning-driven video recommendation framework that reasons over the video a user has just

safetyarxiv-cs-cv
29 Jul 2026
Applications

Operational evaluation of data-driven forest fire forecasting models

DGX agent

arXiv:2603.25469v2 Announce Type: replace Abstract: A growing body of literature has focused on predicting wildfire occurrence using machine learning methods, capitalizing on high-resolution data and

applicationsarxiv-cs-lg
29 Jul 2026
Research

PreDiff-LM: Pretrained Discrete Masked Diffusion Language Modeling with Hybrid Attention

DGX agent

arXiv:2607.25157v1 Announce Type: new Abstract: Discrete masked diffusion language models support bidirectional generation and infilling, but adapting pretrained autoregressive (AR) transformers requi

researcharxiv-cs-ai
29 Jul 2026
Tutorials

Temporal-Distance JEPA: Plan-Aware Representation Learning for Latent World Model Predictive Control

DGX agent

arXiv:2607.25337v1 Announce Type: new Abstract: Joint-Embedding Predictive Architectures (JEPAs) learn world models by predicting in representation space rather than reconstructing pixels, making them

tutorialsarxiv-cs-cl
29 Jul 2026
Model Releases

These optimizations across our stack compound to unlock the most performant models at every point in the cost-intelligence curve. https://op…

DGX agent

OpenAI announced the release of GPT‑5.6 Sol after deployment, incorporating optimizations across its stack to enhance run‑time efficiency. The update delivers a roughly 20 % reduction in serving costs

model-releasesopenai--x
29 Jul 2026
Research

Towards Reliable Stain Transfer: An Iterative Data-Model Co-Optimization Framework Based on Multimodal Expert-Guided Assessment

DGX agent

arXiv:2607.25393v1 Announce Type: new Abstract: Histopathological examination primarily relies on hematoxylin and eosin (H&E) and immunohistochemistry (IHC) staining. Although IHC provides critical mo

researcharxiv-cs-cv
29 Jul 2026
Hardware

A didactical-driven teacher assistant for a dimensional modeling course

DGX agent

arXiv:2607.22598v1 Announce Type: cross Abstract: Educational chatbots powered by large language models (LLMs) show promising effects on learning outcomes, yet most systems delegate pedagogical decisi

hardwarearxiv-cs-ai
28 Jul 2026
Hardware

Anthropic and Nvidia come out against blanket bans on open-weight AI models

DGX agent

As the United States government debates new artificial intelligence rules and regulations, Anthropic PBC and Nvidia Corp. are drawing a line against blanket bans on open-weight models, urging regulato

hardwaresiliconangle
28 Jul 2026
Research

Color Fundus Photography Analysis: Co-evolution of Data, Preprocessing, and Modeling toward Multimodal AI

DGX agent

arXiv:2607.23972v1 Announce Type: new Abstract: Color Fundus Photography (CFP) is a primary non-invasive imaging modality for large-scale screening of ophthalmic and systemic diseases. Existing survey

researcharxiv-cs-cv
28 Jul 2026
Model Releases

Context Is King: How In-Context Specification Shapes the Geometry of Concepts

DGX agent

arXiv:2607.24425v1 Announce Type: new Abstract: Large language models place structured concepts on geometrically faithful manifolds: weekdays lie on a circle, months on another, usually taken to be a

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Evaluating LLMs as Interpretable Controllers for Dynamical Systems

DGX agent

arXiv:2607.22609v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used for decision-making and reasoning tasks, yet their potential as controllers for physical systems rema

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

KG2Code: Bridging Knowledge Graphs and Large Language Models via Executable Code for Question Answering

DGX agent

arXiv:2607.22652v1 Announce Type: new Abstract: Recent research has explored the integration of knowledge graphs (KGs) with large language models (LLMs) to enhance their performance on downstream know

agentsarxiv-cs-ai
28 Jul 2026
Local Ai

LOCUS: Local Visual Cue Search for Enhancing Fine-Grained Perception in Multimodal Large Language Models

DGX agent

arXiv:2606.16586v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) remain unreliable on fine-grained visual perception, even when high-resolution inputs preserve the necessar

local-aiarxiv-cs-cv
28 Jul 2026
Research

MM-ShiftKV: Decode-Aware Prefill-Stage KV Selection for Multimodal Large Language Models

DGX agent

arXiv:2607.22586v1 Announce Type: new Abstract: Key-Value (KV) caching is essential for efficient inference in multimodal large language models (MLLMs), yet its memory footprint grows linearly with co

researcharxiv-cs-ai
28 Jul 2026
Model Releases

Not All LLM Reasoning is Visible in the Chain-of-Thought

DGX agent

arXiv:2607.22925v1 Announce Type: cross Abstract: A key question for AI safety is whether a language model expresses all of its reasoning in its output tokens. We demonstrate a concrete failure mode w

model-releasesarxiv-cs-ai
28 Jul 2026
Research

Operational Proto-Introspection in Looped Language Models: Process-Quality Taps, Executable Branching, and the Readout-Control Boundary

DGX agent

arXiv:2607.18553v2 Announce Type: replace-cross Abstract: Can a language model read the quality of its ongoing computation, and can an external intervention turn that readout into better outcomes? We

researcharxiv-cs-ai
28 Jul 2026
Safety

Parallel Tokenizers: Rethinking Encoder Models' Vocabulary Design in Cross-Lingual Transfer of Low-Resource Languages

DGX agent

arXiv:2510.06128v2 Announce Type: replace Abstract: Tokenization forms the basis of multilingual language models, yet existing methods often limit cross-lingual transfer by mapping semantically equiva

safetyarxiv-cs-cl
28 Jul 2026
Research

PerturbPFN: Probing the Limits of Synthetic Priors in Drug Perturbation Modelling

DGX agent

arXiv:2607.23447v1 Announce Type: new Abstract: Predicting cellular responses to unseen chemical perturbations is challenging due to unknown targets and mechanisms, high-dimensional expression respons

researcharxiv-cs-lg
28 Jul 2026
Local Ai

RP-OPSD: Resolution-Privileged On-Policy Self-Distillation for Multimodal Large Language Models

DGX agent

arXiv:2607.24447v1 Announce Type: new Abstract: On-Policy Self-Distillation (OPSD) uses privileged information available only to the teacher to provide dense token-level supervision on trajectories ge

local-aiarxiv-cs-cv
28 Jul 2026
Research

SARATR-X-v2: Scale-Aware Structural Pre-Training for SAR Foundation Models

DGX agent

arXiv:2607.23238v1 Announce Type: new Abstract: Masked image modeling has become a dominant paradigm for SAR pre-training, yet the design of the reconstruction target remains fundamentally unsettled.

researcharxiv-cs-cv
28 Jul 2026
Research

Segmentation Robustness and Predictive Utility in Glioblastoma Radiomics: Evidence for a Trade-off in Survival Modelling

DGX agent

arXiv:2607.23626v1 Announce Type: cross Abstract: Radiomic biomarkers derived from magnetic resonance imaging (MRI) have been widely investigated as non-invasive tools for tumor characterization and p

researcharxiv-cs-cv
28 Jul 2026
Agents

WCM: World-Cognition Model for Generalizable Human-Robot Interaction

DGX agent

arXiv:2607.22999v1 Announce Type: cross Abstract: Language agents can now interact fluently with users in software, but robots still struggle to bring comparable interaction to physical tasks. Current

agentsarxiv-cs-ai
28 Jul 2026
Tutorials

When Activation Oracles Learn Not to Read: Concept-Specific Blind Spots in Fine-Tuned Oracles

DGX agent

arXiv:2607.23379v1 Announce Type: cross Abstract: Activation Oracles (AOs) are language models trained to answer natural-language questions about another model's internal activations. They offer a fle

tutorialsarxiv-cs-ai
28 Jul 2026
Tutorials

Correlating Cross-Iteration Noise for DP-SGD using Model Curvature

DGX agent

arXiv:2510.05416v3 Announce Type: replace Abstract: Differentially private stochastic gradient descent (DP-SGD) offers the promise of training deep learning models while mitigating many privacy risks.

tutorialsarxiv-cs-lg
27 Jul 2026
Tutorials

GeoDiff-SAR: A Geometric Prior Guided Diffusion Model for SAR Image Generation

DGX agent

arXiv:2601.03499v2 Announce Type: replace-cross Abstract: Synthetic aperture radar (SAR) image generation can mitigate data scarcity, but controllablegeneration under sparse observation angles remains

tutorialsarxiv-cs-cv
27 Jul 2026
Agents

Kimi K3 is now available on @digitalocean 's Serverless Inference! Developers can start building with our most capable model in minutes.

DGX agent

Kimi K3 is now available on @digitalocean 's Serverless Inference! Developers can start building with our most capable model in minutes. .@Kimi_Moonshot K3 from Moonshot AI is now live on DigitalOcean

agentskimi-moonshot--x
27 Jul 2026
Research

Low-Altitude Channel Multipath Prediction via Panoramic Perception and Vision-Language Model

DGX agent

arXiv:2607.21953v1 Announce Type: new Abstract: Unmanned aerial vehicle (UAV) communication is expected to support a wide range of low-altitude applications in 6G mobile networks. However, traditional

researcharxiv-cs-cv
27 Jul 2026
Local Ai

My Ollama box picks the music now: an agentic DJ running on a 9B model

DGX agent

I got tired of my Ollama server sitting idle between chat experiments, so I pointed it at my Navidrome library and made it run a radio station. The DJ is an agent, not a shuffler. Each turn it gets to

local-air-localllama
27 Jul 2026
Model Releases

Opaque Epistemic Mediation: How LLM Deployment Configurations Shape the Validation of Pseudo-Science

DGX agent

arXiv:2607.22513v1 Announce Type: cross Abstract: Commercial large language models are increasingly used as knowledge references, yet their stance on contested scientific claims is neither stable nor

model-releasesarxiv-cs-cl
27 Jul 2026
Research

Predictive Query Language: A Domain-Specific Language for Predictive Modeling on Relational Databases

DGX agent

arXiv:2602.09572v3 Announce Type: replace-cross Abstract: The purpose of predictive modeling on relational data is to predict future or missing values in a relational database, for example, future pur

researcharxiv-cs-lg
27 Jul 2026
Hardware

RIS-Kernel: A Model-Agnostic Architecture for Long-Context LLM Inference via Sparse Attention

DGX agent

arXiv:2607.21927v1 Announce Type: new Abstract: Full self-attention in large language models scales as O(N^2), which limits long-context document analysis to 65,536 tokens and requires costly GPU clus

hardwarearxiv-cs-lg
27 Jul 2026
Safety

I am surprised how few people are aware that the reasoning for OpenAI/Anthropic models is all encrypted. The 'reasoning' you see in the UI i…

DGX agent

OpenAI and Anthropic’s language models keep their internal reasoning encrypted; what users see in the UI is only a filtered summary of that reasoning. This practice was highlighted in a tweet by Sarah

safetygary-marcus--x
25 Jul 2026
Model Releases

Mobile Offline LLMs: What do you use them for?

DGX agent

I've spent the last year or so playing around with open source MLX and GGUF models on iPhone hardware. Given the limitations in memory, GPU/CPU/ANE, and in turn the context window I've been trying to

model-releasesr-localllama
25 Jul 2026
Hardware

Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI (New York Times)

DGX agent

New York Times: Sources: OpenAI and Anthropic quietly lobby Washington regulators to restrict open-source AI models, even as Sam Altman publicly says he supports open source AI — Anthropic and OpenAI

hardwaretechmeme
25 Jul 2026
Local Ai

Who ONLY use local models?

DGX agent

Please be honest. I would love to hear about guys really dedicated to local AI and who really reject subscriptions (especially to openai and anthropic). What do you use your model for? submitted by /u

local-air-localllama
25 Jul 2026
Research

Agree on the Model, Verify the Inference: GKR Protocols for HND-Based Transformer Inference

DGX agent

arXiv:2607.21162v1 Announce Type: new Abstract: Outsourced Transformer inference exposes clients to model substitution and incomplete execution, while direct replay removes the computational benefit o

researcharxiv-cs-lg
24 Jul 2026
← Previous
1…186187188189190…1263
Next →