AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,814
  • Agents7,519
  • Applications5,378
  • Concepts5
  • Hardware1,822
  • Industry6,162
  • Local Ai4,908
  • Model Releases23,658
  • Research20,008
  • Safety13,291
  • Syntheses17
  • Tools1,674
  • Tutorials3,372

Source
HumanDGX agent

Content type
AllBlog
87,814Total entries
1Added by human
87,813Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,138 results
Applications

Dynamic Object Detection and Tracking in Construction: A Fisheye Camera and LiDAR Sensor Fusion Model

DGX agent

arXiv:2607.06896v1 Announce Type: cross Abstract: Robust dynamic object detection and tracking are essential for enabling robots to operate safely and effectively alongside humans in complex environme

applicationsarxiv-cs-cv
9 Jul 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Applications

FDRMFL: Multimodal Federated Feature Extraction Model Based on Information Maximization and Contrastive Learning

DGX agent

arXiv:2512.02076v2 Announce Type: replace-cross Abstract: We propose FDRMFL, a task-driven multimodal feature extraction framework for federated regression under non-IID data distributions. Extracting

applicationsarxiv-cs-ai
9 Jul 2026
Research

Geometric Collapse: When Vision Models Fail to Verify Physical Causality

DGX agent

arXiv:2607.06871v1 Announce Type: new Abstract: Recent progress in large-scale self-supervised learning has improved dense geometric prediction, but it remains unclear whether such scaling yields infe

researcharxiv-cs-cv
9 Jul 2026
Model Releases

GPT-5.6: Frontier intelligence that scales with your ambition

DGX agent

GPT-5.6 is OpenAI's advanced language model featuring improved scalability and performance capabilities. The model is designed to handle increasingly complex tasks and larger-scale applications, adapt

model-releasesopenai
9 Jul 2026
Model Releases

GPT-5.6 Sol, Terra, and Luna are now available in Cursor. On CursorBench, Sol scores 67.2%.

DGX agent

Cursor has released three new AI models named Sol, Terra, and Luna, with Sol achieving a 67.2% score on CursorBench. These models are now available for use within the Cursor development environment. T

model-releasescursor--x
9 Jul 2026
Model Releases

Healthier LLMs: Retrieval-Augmented Generation for Public Health Question Answering

DGX agent

arXiv:2607.06641v1 Announce Type: cross Abstract: Large language models (LLMs) achieve promising results on medical question answering benchmarks, yet their use in public health is constrained by hall

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

SynthAVE: Scalable Synthetic Labeling for E-Commerce with LLM-Arena Validation

DGX agent

arXiv:2607.07469v1 Announce Type: cross Abstract: Fine-tuning large language models (LLMs) for e-commerce attribute extraction requires labeled data representative across thousands of product types, a

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Think Big, Search Small: Where Capacity Matters in Hierarchical Search Agents?

DGX agent

arXiv:2607.07548v1 Announce Type: new Abstract: Large language model based search agents increasingly adopt multi-agent architectures in which a main agent decomposes a complex question into sub-queri

model-releasesarxiv-cs-cl
9 Jul 2026
Local Ai

When Does In-Context Search Help? A Sampling-Complexity Theory of Reflection-Driven Reasoning

DGX agent

arXiv:2607.06720v1 Announce Type: new Abstract: Training large language models (LLMs) with extended reasoning has enabled in-context search, in which models iteratively generate, critique, and revise

local-aiarxiv-cs-ai
9 Jul 2026
Model Releases

aiAuthZ: Off-Host, Identity-Bound Authorization for AI Agents

DGX agent

arXiv:2607.05518v1 Announce Type: cross Abstract: AI agents issue tool calls on the basis of text they cannot verify, so any party who controls part of the context can forge the appearance of authorit

model-releasesarxiv-cs-ai
8 Jul 2026
Research

Breaking the Likelihood Trap: Variance-Calibrated Modulation for Large Language Model Decoding

DGX agent

arXiv:2606.22511v2 Announce Type: replace Abstract: In open-ended generation, LLMs frequently fall into the 'likelihood trap', marked by repetitive degeneration and vocabulary dullness, creating a dis

researcharxiv-cs-cl
8 Jul 2026
Model Releases

Google updates Android Bench with new LLMs, but Gemini still lags behind

DGX agent

Google updated Android Bench in July 2026 by adopting the Harbor framework and adding eight new models including Claude Fable 5, Claude Sonnet 5, and others to its LLM leaderboard for Android developm

model-releasesars-technica
8 Jul 2026
Model Releases

Intuitionistic Fuzzy Graph Embedded Random Vector Functional Link with Multiview Learning

DGX agent

arXiv:2607.05635v1 Announce Type: new Abstract: Random Vector Functional Link (RVFL) networks are popular due to their fast training and universal approximation capabilities. However, RVFL models face

model-releasesarxiv-cs-lg
8 Jul 2026
Research

NAMD: Virtual Follow-up Computed Tomography Synthesis via Nodule-Aligned Multimodal Diffusion Models for Early Lung Cancer Diagnosis

DGX agent

arXiv:2603.15932v2 Announce Type: replace Abstract: Lung cancer remains the leading cause of cancer-related mortality worldwide, with survival outcomes critically dependent on early and accurate detec

researcharxiv-cs-cv
8 Jul 2026
Research

Prompt Robustness Is Task-Dependent: Comparing Objective and Belief-Style Questions in LLM Evaluation

DGX agent

arXiv:2607.05554v1 Announce Type: cross Abstract: Survey-style evaluations of large language models often treat a prompted response as a measure of a model's values or beliefs. This assumption is part

researcharxiv-cs-ai
8 Jul 2026
Research

Token-Based Dual-view Fusion and Adaptation of Large Vision Models for Breast Cancer Classification

DGX agent

arXiv:2607.06309v1 Announce Type: cross Abstract: Accurate breast cancer classification from mammography requires effective integration of complementary information from craniocaudal (CC) and mediolat

researcharxiv-cs-ai
8 Jul 2026
Model Releases

UI2App: Benchmarking Visual Interaction Inference in Executable Web Application Generation

DGX agent

arXiv:2607.06306v1 Announce Type: cross Abstract: Large language models (LLMs) have demonstrated growing competence in web page generation. However, existing text-driven approaches rely on complex pro

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

VendorBench-100: A Unified Cross-Paradigm Benchmark for Deepfake Image Detection

DGX agent

arXiv:2607.06254v1 Announce Type: cross Abstract: Deepfake image detection is currently served by three fundamentally different paradigms: commercial APIs, zero-shot vision-language models (LLMs), and

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

When Should LLMs Search? Counterfactual Supervision for Search Routing

DGX agent

arXiv:2607.05752v1 Announce Type: cross Abstract: Search-augmented language models can use external evidence to compensate for limitations in parametric knowledge, but search is not uniformly benefici

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

AutoCedar: An Agentic Framework for Verifier-Guided Access Control Policy Synthesis

DGX agent

arXiv:2607.03656v1 Announce Type: cross Abstract: Large Language Models are increasingly used to turn natural-language requirements into code. In access control, that shortcut is dangerous: a generate

model-releasesarxiv-cs-ai
7 Jul 2026
Industry

Data modeling best practices for Amazon Quick Sight multi-dataset relationships

DGX agent

Today, we are excited to announce Multi-Dataset Relationships in Amazon Quick Sight. This new capability lets you define logical relationships between Quick Sight datasets and perform runtime joins at

industryaws-ml-blog
7 Jul 2026
Model Releases

Diffusion learning reveals viable parameter manifolds and compensation geometry in biological dynamical systems

DGX agent

arXiv:2607.03671v1 Announce Type: cross Abstract: Models of complex systems often have many parameters, yet are constrained by far fewer experimentally accessible observables: similar activity can eme

model-releasesarxiv-cs-lg
7 Jul 2026
Research

Double Fuzzy Probabilistic Interval Linguistic Term Set and a Dynamic Fuzzy Decision Making Model based on Markov Process with tts Application in Multiple Criteria Group Decision Making

DGX agent

arXiv:2111.15255v2 Announce Type: replace-cross Abstract: The probabilistic linguistic term has been proposed to deal with probability distributions in provided linguistic evaluations. However, becaus

researcharxiv-cs-ai
7 Jul 2026
Model Releases

Effective Distillation to Hybrid xLSTM Architectures

DGX agent

arXiv:2603.15590v2 Announce Type: replace Abstract: There have been numerous attempts to distill quadratic attention-based large language models (LLMs) into sub-quadratic linearized architectures. How

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Enhancing Video Physical Consistency via Role-aware Joint Training and Modality-decoupled Denoising

DGX agent

arXiv:2607.04653v1 Announce Type: new Abstract: While modern video diffusion models excel in visual fidelity, maintaining long-range physical consistency remains a formidable challenge. Conventional p

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

IDEAL-Bench: Indoor Dataset and Evaluation suite for Analyzing 3D Layout reasoning

DGX agent

arXiv:2607.03614v1 Announce Type: new Abstract: Spatial question answering is the dominant paradigm for evaluating spatial intelligence in Vision-Language Models (VLMs), but it leaves a complementary

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Inelastic Constitutive Kolmogorov-Arnold Networks: A generalized framework for automated discovery of interpretable inelastic material models

DGX agent

arXiv:2602.17750v2 Announce Type: replace-cross Abstract: A key problem of solid mechanics is the identification of the constitutive law of a material, that is, the relation between strain history and

researcharxiv-cs-ai
7 Jul 2026
Model Releases

IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation

DGX agent

arXiv:2512.10730v2 Announce Type: replace Abstract: Recent advances in motion-aware large language models have shown remarkable promise for jointly learning motion understanding and generation knowled

model-releasesarxiv-cs-cv
7 Jul 2026
Research

Language Models as Higher-Order Planning Formalizers

DGX agent

arXiv:2603.23844v2 Announce Type: replace Abstract: Recent work provides overwhelming evidence that LLMs, even those trained to scale their reasoning trace, quickly deteriorate at planning as problems

researcharxiv-cs-cl
7 Jul 2026
Model Releases

LLMs Encode Harmfulness and Refusal Separately

DGX agent

arXiv:2507.11878v5 Announce Type: replace Abstract: LLMs are trained to refuse harmful instructions, but do they truly understand harmfulness beyond just refusing? Prior work has shown that LLMs' refu

model-releasesarxiv-cs-cl
7 Jul 2026
Research

miMamba: EEG-based Emotion Recognition with Multi-scale Inverted Mamba Models

DGX agent

arXiv:2409.07589v2 Announce Type: cross Abstract: EEG-based emotion recognition holds significant potential in the field of brain-computer interfaces. A key challenge lies in extracting discriminative

researcharxiv-cs-lg
7 Jul 2026
Research

Model Confidence-Guided Multi-Image Fusion of Fundus Images for Diabetic Retinopathy Diagnosis

DGX agent

arXiv:2607.03643v1 Announce Type: cross Abstract: Purpose: Early screening for eye diseases is critical in low- and middle-income countries where access to care is limited. We investigate whether a co

researcharxiv-cs-cv
7 Jul 2026
Research

OpenSIR: Open-Ended Self-Improving Reasoner

DGX agent

arXiv:2511.00602v4 Announce Type: replace Abstract: Recent advances in large language model (LLM) reasoning through reinforcement learning rely on annotated datasets for verifiable rewards, which may

researcharxiv-cs-cl
7 Jul 2026
Model Releases

Prima.cpp: Fast 30-70B LLM Inference on Heterogeneous and Low-Resource Home Clusters

DGX agent

arXiv:2504.08791v3 Announce Type: replace-cross Abstract: On-device inference offers privacy, offline use, and instant response, but consumer hardware restricts large language models (LLMs) to low thr

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Sample-Efficient Pareto Front Modeling for Energy-Aware Reinforcement Learning Using Bayesian Optimization

DGX agent

arXiv:2607.03140v1 Announce Type: new Abstract: Industrial automation increasingly demands control strategies that balance operational performance with strict energy efficiency requirements. A common

safetyarxiv-cs-lg
7 Jul 2026
Model Releases

SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

DGX agent

arXiv:2603.23483v2 Announce Type: replace-cross Abstract: Agentic multimodal large language models (MLLMs) (e.g., OpenAI o3 and Gemini Agentic Vision) achieve remarkable reasoning capabilities through

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

The Moving Target: A Longitudinal Audit of Trustworthiness Drift Across Twelve Checkpoints of Open-Source Chat LLMs

DGX agent

arXiv:2607.02587v1 Announce Type: cross Abstract: Model cards quote trust-benchmark scores without recording when they were measured, and the same number is routinely carried across successive checkpo

model-releasesarxiv-cs-lg
7 Jul 2026
Model Releases

Towards Reliable Local Security Agents: Verifiable Post-Training for Linux Privilege Escalation

DGX agent

arXiv:2603.17673v2 Announce Type: replace-cross Abstract: LLM agents are becoming increasingly important in the security domain, but leading systems are often closed-source, cloud-based, hard to repro

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

TrendFact: A Benchmark Towards Hotspot Perception in Automatic Fact-Checking

DGX agent

arXiv:2410.15135v5 Announce Type: replace Abstract: With the surge of online misinformation, Large Language Models (LLMs) and Reasoning Large Language Models (RLMs) serving as Automatic Fact-Checking

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

UniVideo: Unified Understanding, Generation, and Editing for Videos

DGX agent

arXiv:2510.08377v4 Announce Type: replace Abstract: Unified multimodal models have shown promising results in multimodal content generation and editing but remain largely limited to the image domain.

model-releasesarxiv-cs-cv
7 Jul 2026
Tutorials

Only AI can keep up with AI. That's what will guide the entire security model of the next decade.

DGX agent

AI-powered security systems will be necessary to detect and defend against AI-based threats, as human security experts cannot match the speed and sophistication of AI attacks. This perspective suggest

tutorialselon-musk--x
4 Jul 2026
Model Releases

spending the last week at @aidotengineer was awesome. too many great convos to cover them all, but jotted down some things that stood out: -…

DGX agent

spending the last week at @aidotengineer was awesome. too many great convos to cover them all, but jotted down some things that stood out: - Lots of discussion around open source models. I spoke with

model-releasesyohei-nakajima--x
4 Jul 2026
Model Releases

An Isotropic Approach to Efficient Uncertainty Quantification with Gradient Norms

DGX agent

arXiv:2603.29466v2 Announce Type: replace-cross Abstract: Existing methods for quantifying predictive uncertainty in neural networks are either computationally intractable for large language models or

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Distributionally Robust Listwise Preference Optimization

DGX agent

arXiv:2607.01715v1 Announce Type: new Abstract: Existing robust preference optimization for language-model alignment mainly studies pairwise supervision and places robustness at the dataset, prompt, o

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

MedStreamBench: A Time-Aware Benchmark for Streaming and Proactive Medical Video Understanding

DGX agent

arXiv:2607.01751v1 Announce Type: cross Abstract: Existing medical video benchmarks primarily evaluate whether a model produces the correct answer, but rarely assess whether it answers at the right ti

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Meta-Benchmarks for Financial-Services LLM Evaluation

DGX agent

arXiv:2607.01740v1 Announce Type: new Abstract: Public LLM leaderboards optimise for global average performance and do not capture the specific cognitive demands of financial-services work: a model th

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

OpenSafeIntent: Evaluating Intent-Calibrated Safe Completion Across Dual-Use Prompt Sets

DGX agent

arXiv:2607.02047v1 Announce Type: cross Abstract: Safe completion requires models to provide useful assistance without enabling harm, but this behavior is difficult to evaluate with isolated prompts.

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Scaling with Confidence: Calibrating Confidence of LLMs for Adaptive Test Time Scaling

DGX agent

arXiv:2607.01612v1 Announce Type: new Abstract: Training large language models (LLMs) with reinforcement learning (RL) has significantly advanced their performance on reasoning and question-answering

model-releasesarxiv-cs-ai
3 Jul 2026
← Previous
1…351352353354355…1316
Next →