AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Research

Structured Abductive-Deductive-Inductive Reasoning for LLMs via Algebraic Invariants

DGX agent

arXiv:2604.15727v1 Announce Type: new Abstract: Large language models exhibit systematic limitations in structured logical reasoning: they conflate hypothesis generation with verification, cannot dist

researcharxiv-cs-ai
20 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Subjective and Objective Quality-of-Experience Evaluation Study for Live Video Streaming

DGX agent

arXiv:2409.17596v2 Announce Type: replace-cross Abstract: In recent years, live video streaming has gained widespread popularity across various social media platforms. Quality of experience (QoE), whi

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination

DGX agent

arXiv:2510.22977v2 Announce Type: replace-cross Abstract: Enhancing the reasoning capabilities of Large Language Models (LLMs) is a key strategy for building Agents that 'think then act.' However, rec

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Towards Rigorous Explainability by Feature Attribution

DGX agent

arXiv:2604.15898v1 Announce Type: new Abstract: For around a decade, non-symbolic methods have been the option of choice when explaining complex machine learning (ML) models. Unfortunately, such metho

researcharxiv-cs-ai
20 Apr 2026
Research

Training Flow Matching: The Role of Weighting and Parameterization

DGX agent

arXiv:2603.06454v2 Announce Type: replace Abstract: We study the training objectives of denoising-based generative models, with a particular focus on loss weighting and output parameterization, includ

researcharxiv-cs-cv
20 Apr 2026
Model Releases

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions

DGX agent

arXiv:2604.15823v1 Announce Type: new Abstract: Embodied robotic agents often perceive movies through an egocentric screen-view interface rather than native cinematic footage, introducing domain shift

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

WildFeedback: Aligning LLMs With In-situ User Interactions And Feedback

DGX agent

arXiv:2408.15549v4 Announce Type: replace Abstract: As large language models (LLMs) continue to advance, aligning these models with human preferences has emerged as a critical challenge. Traditional a

safetyarxiv-cs-cl
20 Apr 2026
Research

Zoom Consistency: A Free Confidence Signal in Multi-Step Visual Grounding Pipelines

DGX agent

arXiv:2604.15376v1 Announce Type: cross Abstract: Multi-step zoom-in pipelines are widely used for GUI grounding, yet the intermediate predictions they produce are typically discarded after coordinate

researcharxiv-cs-ai
20 Apr 2026
Research

“Salaryman eating ramen” is like the Eastern equivalent of “Will Smith eating spaghetti” test

DGX agent

This post draws a humorous parallel between using images of salarymen eating ramen as a test case for AI image generation models in Eastern contexts and the Western 'Will Smith eating spaghetti' meme,

researchdavid-ha--x
19 Apr 2026
Model Releases

Since Anthropic publish their system prompts we can generate a diff between Claude Opus 4.6 and 4.7 - here are my notes on what's changed ht…

DGX agent

Simon Willison documents the differences between Anthropic's Claude Opus 4.6 and 4.7 system prompts, analyzing changes that Anthropic made public. The notes likely highlight modifications to model beh

model-releasessimon-willison--x
19 Apr 2026
Local Ai

Sweet spot…Cloud & local LLM setup + Mission Control

DGX agent

A discussion exploring the 'sweet spot' of using Ollama's hybrid Cloud + Local setup, where a reachable Ollama host serves as the control point for both local and cloud models . The post likely covers

local-air-ollama
18 Apr 2026
Model Releases

3D Instruction Ambiguity Detection

DGX agent

arXiv:2601.05991v2 Announce Type: replace Abstract: In safety-critical domains, linguistic ambiguity can have severe consequences; a vague command like 'Pass me the vial' in a surgical setting could l

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Active Learning with Selective Time-Step Acquisition for PDEs

DGX agent

arXiv:2511.18107v2 Announce Type: replace Abstract: Accurately solving partial differential equations (PDEs) is critical to understanding complex scientific and engineering phenomena, yet traditional

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

AI-Assisted Peer Review at Scale: The AAAI-26 AI Review Pilot

DGX agent

arXiv:2604.13940v1 Announce Type: new Abstract: Scientific peer review faces mounting strain as submission volumes surge, making it increasingly difficult to sustain review quality, consistency, and t

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

[AINews] Anthropic Claude Opus 4.7 - literally one step better than 4.6 in every dimension

DGX agent

This article from Latent Space discusses Anthropic's Claude Opus 4.7 release, highlighting incremental improvements across multiple performance dimensions compared to the previous 4.6 version. The pie

model-releaseslatent-space
17 Apr 2026
Model Releases

Applying an Agentic Coding Tool for Improving Published Algorithm Implementations

DGX agent

arXiv:2604.13109v1 Announce Type: cross Abstract: We present a two-stage pipeline for AI-assisted improvement of published algorithm implementations. In the first stage, a large language model with re

model-releasesarxiv-cs-ai
17 Apr 2026
Research

ArrowGEV: Grounding Events in Video via Learning the Arrow of Time

DGX agent

arXiv:2601.06559v2 Announce Type: replace Abstract: Grounding events in videos serves as a fundamental capability in video analysis. While Vision Language Models (VLMs) are increasingly employed for t

researcharxiv-cs-cv
17 Apr 2026
Model Releases

Assessment Design in the AI Era: A Method for Identifying Items Functioning Differentially for Humans and Chatbots

DGX agent

arXiv:2603.23682v2 Announce Type: replace-cross Abstract: The rapid adoption of large language models (LLMs) in education raises profound challenges for assessment design. To adapt assessments to the

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

ChatSVA: Bridging SVA Generation for Hardware Verification via Task-Specific LLMs

DGX agent

arXiv:2604.02811v2 Announce Type: replace-cross Abstract: Functional verification consumes over 50% of the IC development lifecycle, where SystemVerilog Assertions (SVAs) are indispensable for formal

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Class Unlearning via Depth-Aware Removal of Forget-Specific Directions

DGX agent

arXiv:2604.15166v1 Announce Type: new Abstract: Machine unlearning aims to remove targeted knowledge from a trained model without the cost of retraining from scratch. In class unlearning, however, red

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

Comparison of Modern Multilingual Text Embedding Techniques for Hate Speech Detection Task

DGX agent

arXiv:2604.14907v1 Announce Type: new Abstract: Online hate speech and abusive language pose a growing challenge for content moderation, especially in multilingual settings and for low-resource langua

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Create Expert Content: Deploying a Multi-Agent System with Terraform and Cloud Run

DGX agent

In support of our mission to accelerate the developer journey on Google Cloud, we built Dev Signal: a multi-agent system designed to transform raw community signals into reliable technical guidance by

model-releasesgoogle-cloud-ai
17 Apr 2026
Local Ai

DocVAL: Validated Chain-of-Thought Distillation for Grounded Document VQA

DGX agent

arXiv:2511.22521v2 Announce Type: replace Abstract: Document visual question answering requires models not only to answer questions correctly, but also to precisely localize answers within complex doc

local-aiarxiv-cs-cv
17 Apr 2026
Model Releases

Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG

DGX agent

arXiv:2604.14572v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds LLM responses in external evidence but treats the model as a passive consumer of search results: it never

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

Evolving Media CDN for the world’s most demanding broadcast and streaming workloads

DGX agent

Editor’s note: In this post, we share joint insights from Raj Gulani, Director of Product Management for Network Experiences, and Dan Rayburn, Industry analyst with 30-plus years of experience coverin

model-releasesgoogle-cloud-ai
17 Apr 2026
Research

Federated Breast Cancer Detection Enhanced by Synthetic Ultrasound Image Augmentation

DGX agent

arXiv:2506.23334v3 Announce Type: replace-cross Abstract: Federated learning enables collaborative training of deep learning models across institutions without sharing sensitive patient data. However,

researcharxiv-cs-cv
17 Apr 2026
Tutorials

Federated Multi-Task Clustering

DGX agent

arXiv:2512.22897v3 Announce Type: replace Abstract: Spectral clustering has emerged as one of the most effective clustering algorithms due to its superior performance. However, most existing models ar

tutorialsarxiv-cs-lg
17 Apr 2026
Model Releases

From Memorization to Creativity: LLM as a Designer of Novel Neural Architectures

DGX agent

arXiv:2601.02997v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel in program synthesis, yet their capacity for neural architecture design -- balancing syntactic reliability,

model-releasesarxiv-cs-cv
17 Apr 2026
Applications

Graph-Based Fraud Detection with Dual-Path Graph Filtering

DGX agent

arXiv:2604.14235v1 Announce Type: new Abstract: Fraud detection on graph data can be viewed as a demanding task that requires distinguishing between different types of nodes. Because graph neural netw

applicationsarxiv-cs-lg
17 Apr 2026
Safety

IG-Search: Step-Level Information Gain Rewards for Search-Augmented Reasoning

DGX agent

arXiv:2604.15148v1 Announce Type: cross Abstract: Reinforcement learning has emerged as an effective paradigm for training large language models to perform search-augmented reasoning. However, existin

safetyarxiv-cs-cl
17 Apr 2026
Safety

Learning Adaptive Reasoning Paths for Efficient Visual Reasoning

DGX agent

arXiv:2604.14568v1 Announce Type: cross Abstract: Visual reasoning models (VRMs) have recently shown strong cross-modal reasoning capabilities by integrating visual perception with language reasoning.

safetyarxiv-cs-cl
17 Apr 2026
Research

Leveraging graph neural networks and mobility data for COVID-19 forecasting

DGX agent

arXiv:2501.11711v2 Announce Type: replace Abstract: The COVID-19 pandemic has claimed millions of lives, spurring the development of diverse forecasting models. In this context, the true utility of co

researcharxiv-cs-lg
17 Apr 2026
Model Releases

LLM Predictive Scoring and Validation: Inferring Experience Ratings from Unstructured Text

DGX agent

arXiv:2604.14321v1 Announce Type: new Abstract: We tasked GPT-4.1 to read what baseball fans wrote about their game-day experience and predict the overall experience rating each fan gave on a 0-10 sur

model-releasesarxiv-cs-cl
17 Apr 2026
Research

MambaSL: Exploring Single-Layer Mamba for Time Series Classification

DGX agent

arXiv:2604.15174v1 Announce Type: new Abstract: Despite recent advances in state space models (SSMs) such as Mamba across various sequence domains, research on their standalone capacity for time serie

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Maximal Brain Damage Without Data or Optimization: Disrupting Neural Networks via Sign-Bit Flips

DGX agent

arXiv:2502.07408v2 Announce Type: replace-cross Abstract: Deep Neural Networks (DNNs) can be catastrophically disrupted by flipping only a handful of parameter bits. We introduce Deep Neural Lesion (D

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

ModuSeg: Decoupling Object Discovery and Semantic Retrieval for Training-Free Weakly Supervised Segmentation

DGX agent

arXiv:2604.07021v2 Announce Type: replace Abstract: Weakly supervised semantic segmentation aims to achieve pixel-level predictions using image-level labels. Existing methods typically entangle semant

model-releasesarxiv-cs-cv
17 Apr 2026
Research

On the Expressive Power and Limitations of Multi-Layer SSMs

DGX agent

arXiv:2604.14501v1 Announce Type: new Abstract: We study the expressive power and limitations of multi-layer state-space models (SSMs). First, we show that multi-layer SSMs face fundamental limitation

researcharxiv-cs-lg
17 Apr 2026
Model Releases

Open-Set Vein Biometric Recognition with Deep Metric Learning

DGX agent

arXiv:2604.14874v1 Announce Type: new Abstract: Most state-of-the-art vein recognition methods rely on closed-set classification, which inherently limits their scalability and prevents the adaptive en

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

Preconditioned Test-Time Adaptation for Out-of-Distribution Debiasing in Narrative Generation

DGX agent

arXiv:2603.13683v2 Announce Type: replace Abstract: Although debiased large language models (LLMs) excel at handling known or low-bias prompts, they often fail on unfamiliar and high-bias prompts. We

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Query pipeline optimization for cancer patient question answering systems

DGX agent

arXiv:2412.14751v2 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) mitigates hallucination in Large Language Models (LLMs) by using query pipelines to retrieve relevant external

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

ReviewGrounder: Improving Review Substantiveness with Rubric-Guided, Tool-Integrated Agents

DGX agent

arXiv:2604.14261v1 Announce Type: new Abstract: The rapid rise in AI conference submissions has driven increasing exploration of large language models (LLMs) for peer review support. However, LLM-base

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Reward-Aware Trajectory Shaping for Few-step Visual Generation

DGX agent

arXiv:2604.14910v1 Announce Type: new Abstract: Achieving high-fidelity generation in extremely few sampling steps has long been a central goal of generative modeling. Existing approaches largely rely

safetyarxiv-cs-cv
17 Apr 2026
Model Releases

SafeHarness: Lifecycle-Integrated Security Architecture for LLM-based Agent Deployment

DGX agent

arXiv:2604.13630v1 Announce Type: cross Abstract: The performance of large language model (LLM) agents depends critically on the execution harness, the system layer that orchestrates tool use, context

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local …

DGX agent

Sharing my current setup to run Qwen3.6 locally in a good agentic setup (Pi + llama.cpp). Should give you a good overview of how good local agents are today: # Start llama.cpp server: llama-server -hf

model-releasesclem-delangue--x
17 Apr 2026
Industry

Tesla is officially launching in Estonia. The company is holding an opening event on April 24th at Ülemiste Center 'Bring your friends and f…

DGX agent

Tesla is officially launching in Estonia. The company is holding an opening event on April 24th at Ülemiste Center 'Bring your friends and family to see our models, take part in the day’s activities a

industryelon-musk--x
17 Apr 2026
Applications

Text2Arch: A Dataset for Generating Scientific Architecture Diagrams from Natural Language Descriptions

DGX agent

arXiv:2604.14941v1 Announce Type: new Abstract: Communicating complex system designs or scientific processes through text alone is inefficient and prone to ambiguity. A system that automatically gener

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

Theory of Mind in Action: The Instruction Inference Task in Dynamic Human-Agent Collaboration

DGX agent

arXiv:2507.02935v2 Announce Type: replace Abstract: Successful human-agent teaming relies on an agent being able to understand instructions given by a (human) principal. In many cases, an instruction

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

Towards AI-assisted Neutrino Flavor Theory Design

DGX agent

arXiv:2506.08080v2 Announce Type: replace-cross Abstract: Particle physics theories, such as those which explain neutrino flavor mixing, arise from a vast landscape of model-building possibilities. A

agentsarxiv-cs-lg
17 Apr 2026
← Previous
1…556557558559560…1371
Next →