AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,457
  • Agents7,399
  • Applications5,302
  • Concepts5
  • Hardware1,786
  • Industry6,117
  • Local Ai4,835
  • Model Releases23,193
  • Research19,715
  • Safety13,094
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,457Total entries
1Added by human
86,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,039 results
16 Jul 2026

Quantum Circuit Vision: Cost-Aware Evaluation of Visual AI Agents for Quantum Code Generation

Model ReleasesDGX agent

arXiv:2607.10057v1 Announce Type: cross Abstract: Can AI agents visually comprehend quantum circuit diagrams and generate verified executable code--and at what cost? We present Quantum Circuit Vision,

15 Jul 2026

AdaPCLA: Adaptive Prior-Calibrated Logit Adjustment for Long-Tailed Longitudinal EHR Generation

ApplicationsDGX agent

arXiv:2607.12645v1 Announce Type: new Abstract: Generative modeling of longitudinal Electronic Health Records is increasingly important for privacy-preserving research, yet standard autoregressive mod

Agentic systems for breast cancer treatment recommendations

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

arXiv:2607.12051v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly being explored for clinical decision support, but their reliability in complex oncology treatment planning

Data Safety: Synthetic Data Quality Analysis Using CIFAKE Dataset

SafetyDGX agent

arXiv:2607.12165v1 Announce Type: new Abstract: Recently, the societal implementation of high-performance image classification models has expanded rapidly. While these models require vast amounts of t

Efficiently Learning Branching Networks for Multitask Algorithmic Reasoning

Model ReleasesDGX agent

arXiv:2512.01113v2 Announce Type: replace-cross Abstract: Algorithmic reasoning -- the ability to perform step-by-step logical inference -- is a synthetic benchmark for evaluating multi-step reasoning

Evaluating Reliability in Machine Learning Models for Early Chronic Kidney Disease Prediction: A Systematic Review of Data Leakage and Predictor Stability

ApplicationsDGX agent

arXiv:2607.11963v1 Announce Type: cross Abstract: The early detection of Chronic Kidney Disease using machine learning has attracted significant interest in healthcare-related computer science. Despit

Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning

Model ReleasesDGX agent

arXiv:2509.24372v3 Announce Type: replace-cross Abstract: Fine-tuning large language models (LLMs) for downstream tasks is an essential stage of modern AI deployment. Reinforcement learning (RL) has e

How Inference Compute Shapes Frontier LLM Evaluation

Model ReleasesDGX agent

arXiv:2606.17930v2 Announce Type: replace Abstract: AI evaluations are shifting toward harder tasks that benefit from longer trajectories involving tool use and iterative problem solving. As a result,

Lightweight Multi-Scale Anomaly Detection for Resource-Constrained Edge Devices

Model ReleasesDGX agent

arXiv:2607.12599v1 Announce Type: new Abstract: Time-series anomaly detection is increasingly important in IoT systems, sensor networks, and edge monitoring applications, where models must operate und

MQAdapter: Multi-Modal Quantum Adapter for Coarse-to-Fine VLM Fine-tuning

Model ReleasesDGX agent

arXiv:2607.12418v1 Announce Type: new Abstract: Large-scale Vision-Language Models have demonstrated impressive transfer learning capabilities across a wide range of tasks. For few-shot classification

Multi-Perspective Agentic Program Repair via Code Property Graphs and Temporal Execution Graphs

Model ReleasesDGX agent

arXiv:2607.12605v1 Announce Type: cross Abstract: Large language models (LLMs) have improved automated program repair (APR), but two limitations remain. First, raw execution traces are often too large

RealSkin: Spatio-Spectral Partial Neural Adjoint Maps for Image-to-3D Attribute Transfer

Model ReleasesDGX agent

arXiv:2607.12495v1 Announce Type: new Abstract: Creating photorealistic 3D assets requires bridging the appearance gap between real-world observations and synthetic models. A promising approach is to

ReDiTT: Retrieval Augmented Conditional Diffusion Transformers for Asynchronous Time Series

ApplicationsDGX agent

arXiv:2607.12391v1 Announce Type: new Abstract: We present a diffusion based model for asynchronous time series prediction, where the goal is to predict the next inter event time and event type. To ad

RFMSR: Residual Flow Matching for Image Super-Resolution

ResearchDGX agent

arXiv:2607.12753v1 Announce Type: new Abstract: Image super-resolution (ISR) has witnessed remarkable progress with diffusion models and flow matching. The dominant text-to-image (T2I) based approache

Translation as a Computationally Efficient Bridge: Feasibility of English BERT for Low-Resource Languages

ResearchDGX agent

arXiv:2607.12612v1 Announce Type: new Abstract: BERT models have revolutionised Natural Language Processing (NLP) through their ability to process unstructured text across diverse domains. However, de

When Directional Accuracy Lies: A Base-Rate-Honest Benchmark for LoRA-Adapted TimesFM on Equity Forecasting

Model ReleasesDGX agent

arXiv:2607.12248v1 Announce Type: cross Abstract: Large pretrained time-series models such as TimesFM are attractive for financial forecasting, but raw directional accuracy is a misleading scoreboard

14 Jul 2026

multiple such incidents have been reported. a clear reminder that current AI cannot be trusted. in racing these techniques ahead, we are ask…

Model ReleasesDGX agent

multiple such incidents have been reported. a clear reminder that current AI cannot be trusted. in racing these techniques ahead, we are asking for trouble GPT-5.6 Sol just deleted my whole production

13 Jul 2026

so on-brand

Model ReleasesDGX agent

so on-brand >be openai >release gpt 5.6 sol >max thinking budgets, full psycho >benchmark results take few days to roll in >“best model ever” >reviews locked in >4 days later >silently nerf it >nobody

10 Jul 2026

A Graph Neural Network Model for Real-Time Gesture Recognition Based on sEMG Signals

ResearchDGX agent

arXiv:2607.07850v1 Announce Type: new Abstract: For seemless control of advanced hand prostheses and augmented reality, accurate and immediate hand gestures recognition is essential. Surface electromy

a hot tip for @GoogleDeepMind : open source gemini 3.5 pro and become the most beloved tech co overnight.

Model ReleasesDGX agent

Clem Delangue suggests that Google DeepMind could significantly improve its public reputation by open-sourcing the Gemini 3.5 Pro model. The post implies that releasing this advanced AI model as open

A Reliability Assessment of LALM Audio Judges for Full-Duplex Voice Agents

Model ReleasesDGX agent

arXiv:2607.07985v1 Announce Type: cross Abstract: We report the empirical reliability of Gemini models as audio judges that score full-duplex agent conversations directly from the raw stereo waveform,

Adaptive Generation of Bias-Eliciting Questions for LLMs

Model ReleasesDGX agent

arXiv:2510.12857v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now widely deployed in user-facing applications, reaching hundreds of millions of users worldwide. Despite th

Adversarial Social Epistemology for Assemblies of Humans and Large Language Models

ResearchDGX agent

arXiv:2607.07760v1 Announce Type: new Abstract: We outline an adversarial social epistemology (ASE) for densely interactive communicative landscapes in which public assertions are scaffolded by chains

Aligning Clinical Needs and AI Capabilities: A Survey on LLMs for Medical Reasoning

Model ReleasesDGX agent

arXiv:2607.07761v1 Announce Type: new Abstract: Large language models (LLMs) have emerged as important tools in healthcare, showing growing potential for clinical reasoning and patient care. This surv

CausalDS: Benchmarking Causal Reasoning in Data-Science Agents

Model ReleasesDGX agent

arXiv:2607.08093v1 Announce Type: new Abstract: Large language models (LLMs) increasingly act as integrated data-science agents, combining abstract reasoning with advanced tool use. Yet the relevant b

Closing the Null Space: Guidance-Aware Quantization for Classifier-Free Diffusion

Model ReleasesDGX agent

arXiv:2607.08241v1 Announce Type: new Abstract: Deploying classifier-free guidance (CFG) diffusion models under real-world compute budgets requires quantization, yet existing post-training quantizatio

Data Alchemy: Mitigating Cross-Site Model Variability Through Test Time Data Calibration

ResearchDGX agent

arXiv:2407.13632v2 Announce Type: replace Abstract: Deploying deep learning-based imaging tools across various clinical sites poses significant challenges due to inherent domain shifts and regulatory

Different Teachers, Different Capabilities: Sub-1B On-Device Distillation for Structured Text Enrichment

Model ReleasesDGX agent

arXiv:2607.08268v1 Announce Type: new Abstract: High-volume structured extraction pays a large model's latency on every item, so distilling the task into a small on-device model is attractive: compara

Federated Deep Learning for Privacy-Preserving Cardiovascular Disease Risk Prediction

ApplicationsDGX agent

arXiv:2607.08595v1 Announce Type: new Abstract: Cardiovascular disease risk prediction models often rely on data from a single institution or centrally pooled datasets. Extending these models across i

FedTR: Federated Learning Framework with Transfer Learning for Industrial Visual Inspection

Local AiDGX agent

arXiv:2607.08014v1 Announce Type: new Abstract: Federated learning (FL) is a collaborative learning scheme to train deep learning models, where collaborating parties can consolidate their models witho

GPT-5.6 Terra and Sol are now available in Perplexity and Perplexity Computer.

Model ReleasesDGX agent

Perplexity has released two new AI models, GPT-5.6 Terra and Sol, which are now available for use in Perplexity's search platform and Perplexity Computer tool. These models represent updates to Perple

Grok doesn’t give up

Model ReleasesDGX agent

Grok doesn’t give up Grok 4.5 is the most persistent agent model we've tested. Here's one example: In one of our evals, we asked 3 models (GPT-5.5, GLM-5.2 and Grok 4.5) to audit a GitHub repo for har

Native Video-Action Pretraining for Generalizable Robot Control

SafetyDGX agent

arXiv:2607.08639v1 Announce Type: cross Abstract: The advent of video-action models offers a promising path for robot control. Nevertheless, we argue that repurposing video generative models designed

Resolving Predictive Multiplicity for the Rashomon Set

Local AiDGX agent

arXiv:2601.09071v2 Announce Type: replace Abstract: The existence of multiple, equally accurate models for a given predictive task leads to predictive multiplicity, where a Rashomon set of models achi

Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning

ResearchDGX agent

arXiv:2607.08393v1 Announce Type: new Abstract: Fine-tuning LLMs to inject new knowledge faces a critical challenge: LLMs can quickly memorize new facts, yet fail to use them for downstream reasoning

Understanding Axes of Difficulty For Long Context Tasks Via PredicateLongBench

Model ReleasesDGX agent

arXiv:2607.08284v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated rapidly improving long-context capabilities, prompting a wave of benchmarks designed to evaluate them. Ho

9 Jul 2026

Bound to Disagree: Generalization Bounds via Certifiable Surrogates

ResearchDGX agent

arXiv:2602.23128v2 Announce Type: replace Abstract: Generalization bounds for deep learning models are typically vacuous, not computable or restricted to specific model classes. In this paper, we tack

Comparative Study of Domain-adapted VLMs for General Document Visual Question Answering

Model ReleasesDGX agent

arXiv:2607.07179v1 Announce Type: new Abstract: Document Visual Question Answering (DocVQA) presents a complex multimodal challenge, requiring models to exploit visual, textual, and layout information

Complexity-Budgeted, Interaction-Aware Interpretable Model for Tabular Data

ApplicationsDGX agent

arXiv:2607.07060v1 Announce Type: cross Abstract: Inherently interpretable classifiers for tabular data typically rely on sparse features, rules, or patterns that users can inspect directly. The margi

Dynamic Object Detection and Tracking in Construction: A Fisheye Camera and LiDAR Sensor Fusion Model

ApplicationsDGX agent

arXiv:2607.06896v1 Announce Type: cross Abstract: Robust dynamic object detection and tracking are essential for enabling robots to operate safely and effectively alongside humans in complex environme

FDRMFL: Multimodal Federated Feature Extraction Model Based on Information Maximization and Contrastive Learning

ApplicationsDGX agent

arXiv:2512.02076v2 Announce Type: replace-cross Abstract: We propose FDRMFL, a task-driven multimodal feature extraction framework for federated regression under non-IID data distributions. Extracting

Geometric Collapse: When Vision Models Fail to Verify Physical Causality

ResearchDGX agent

arXiv:2607.06871v1 Announce Type: new Abstract: Recent progress in large-scale self-supervised learning has improved dense geometric prediction, but it remains unclear whether such scaling yields infe

GPT-5.6: Frontier intelligence that scales with your ambition

Model ReleasesDGX agent

GPT-5.6 is OpenAI's advanced language model featuring improved scalability and performance capabilities. The model is designed to handle increasingly complex tasks and larger-scale applications, adapt

GPT-5.6 Sol, Terra, and Luna are now available in Cursor. On CursorBench, Sol scores 67.2%.

Model ReleasesDGX agent

Cursor has released three new AI models named Sol, Terra, and Luna, with Sol achieving a 67.2% score on CursorBench. These models are now available for use within the Cursor development environment. T

Healthier LLMs: Retrieval-Augmented Generation for Public Health Question Answering

Model ReleasesDGX agent

arXiv:2607.06641v1 Announce Type: cross Abstract: Large language models (LLMs) achieve promising results on medical question answering benchmarks, yet their use in public health is constrained by hall

SynthAVE: Scalable Synthetic Labeling for E-Commerce with LLM-Arena Validation

Model ReleasesDGX agent

arXiv:2607.07469v1 Announce Type: cross Abstract: Fine-tuning large language models (LLMs) for e-commerce attribute extraction requires labeled data representative across thousands of product types, a

Think Big, Search Small: Where Capacity Matters in Hierarchical Search Agents?

Model ReleasesDGX agent

arXiv:2607.07548v1 Announce Type: new Abstract: Large language model based search agents increasingly adopt multi-agent architectures in which a main agent decomposes a complex question into sub-queri

When Does In-Context Search Help? A Sampling-Complexity Theory of Reflection-Driven Reasoning

Local AiDGX agent

arXiv:2607.06720v1 Announce Type: new Abstract: Training large language models (LLMs) with extended reasoning has enabled in-context search, in which models iteratively generate, critique, and revise

8 Jul 2026

aiAuthZ: Off-Host, Identity-Bound Authorization for AI Agents

Model ReleasesDGX agent

arXiv:2607.05518v1 Announce Type: cross Abstract: AI agents issue tool calls on the basis of text they cannot verify, so any party who controls part of the context can forge the appearance of authorit

Breaking the Likelihood Trap: Variance-Calibrated Modulation for Large Language Model Decoding

ResearchDGX agent

arXiv:2606.22511v2 Announce Type: replace Abstract: In open-ended generation, LLMs frequently fall into the 'likelihood trap', marked by repetitive degeneration and vocabulary dullness, creating a dis

Google updates Android Bench with new LLMs, but Gemini still lags behind

Model ReleasesDGX agent

Google updated Android Bench in July 2026 by adopting the Harbor framework and adding eight new models including Claude Fable 5, Claude Sonnet 5, and others to its LLM leaderboard for Android developm

Intuitionistic Fuzzy Graph Embedded Random Vector Functional Link with Multiview Learning

Model ReleasesDGX agent

arXiv:2607.05635v1 Announce Type: new Abstract: Random Vector Functional Link (RVFL) networks are popular due to their fast training and universal approximation capabilities. However, RVFL models face

NAMD: Virtual Follow-up Computed Tomography Synthesis via Nodule-Aligned Multimodal Diffusion Models for Early Lung Cancer Diagnosis

ResearchDGX agent

arXiv:2603.15932v2 Announce Type: replace Abstract: Lung cancer remains the leading cause of cancer-related mortality worldwide, with survival outcomes critically dependent on early and accurate detec

Prompt Robustness Is Task-Dependent: Comparing Objective and Belief-Style Questions in LLM Evaluation

ResearchDGX agent

arXiv:2607.05554v1 Announce Type: cross Abstract: Survey-style evaluations of large language models often treat a prompted response as a measure of a model's values or beliefs. This assumption is part

Token-Based Dual-view Fusion and Adaptation of Large Vision Models for Breast Cancer Classification

ResearchDGX agent

arXiv:2607.06309v1 Announce Type: cross Abstract: Accurate breast cancer classification from mammography requires effective integration of complementary information from craniocaudal (CC) and mediolat

UI2App: Benchmarking Visual Interaction Inference in Executable Web Application Generation

Model ReleasesDGX agent

arXiv:2607.06306v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex pro

VendorBench-100: A Unified Cross-Paradigm Benchmark for Deepfake Image Detection

Model ReleasesDGX agent

arXiv:2607.06254v1 Announce Type: cross Abstract: Deepfake image detection is currently served by three fundamentally different paradigms: commercial APIs, zero-shot vision-language models (LLMs), and

When Should LLMs Search? Counterfactual Supervision for Search Routing

Model ReleasesDGX agent

arXiv:2607.05752v1 Announce Type: cross Abstract: Search-augmented language models can use external evidence to compensate for limitations in parametric knowledge, but search is not uniformly benefici

7 Jul 2026

AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis

Model ReleasesDGX agent

arXiv:2607.03656v1 Announce Type: cross Abstract: Large Language Models are increasingly used to turn natural-language requirements into code. In access control, that shortcut is dangerous: a generate

Data modeling best practices for Amazon Quick Sight multi-dataset relationships

IndustryDGX agent

Today, we are excited to announce Multi-Dataset Relationships in Amazon Quick Sight. This new capability lets you define logical relationships between Quick Sight datasets and perform runtime joins at

← Previous
1…274275276277278…1034
Next →