AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,376
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,170
  • Local Ai4,930
  • Model Releases23,883
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,376Total entries
1Added by human
88,375Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,602 results
Safety

Momentum-Anchored Multi-Scale Fusion Model for Long-Tailed Chest X-Ray Classification

DGX agent

arXiv:2605.02292v1 Announce Type: new Abstract: Chest X-ray classification suffers from severe class imbalance where gradient updates bias toward majority classes, causing feature drift and poor perfo

safetyarxiv-cs-cv
5 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Multi-fidelity surrogates for mechanics of composites: from co-kriging to multi-fidelity neural networks

DGX agent

arXiv:2605.02871v1 Announce Type: cross Abstract: Composite materials exhibit strongly hierarchical and anisotropic properties governed by coupled mechanisms spanning constituents, plies, laminates, s

model-releasesarxiv-cs-lg
5 May 2026
Applications

Multimodal Confidence Modeling in Audio-Visual Quality Assessment

DGX agent

arXiv:2605.01219v1 Announce Type: cross Abstract: Audio-visual quality assessment (AVQA) is essential for streaming, teleconferencing, and immersive media. In realistic streaming scenarios, distortion

applicationsarxiv-cs-cv
5 May 2026
Model Releases

On Stable Long-Form Generation: Benchmarking and Mitigating Length Volatility

DGX agent

arXiv:2605.01357v1 Announce Type: new Abstract: Large Language Models (LLMs) excel at long-context understanding but exhibit significant limitations in long-form generation. Existing studies primarily

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

PepSpecBench: A Unified Evaluation Benchmark for Peptide Tandem Mass Spectrometry Prediction

DGX agent

arXiv:2605.01945v1 Announce Type: new Abstract: Tandem mass spectrometry provides a high-throughput framework for identifying and quantifying proteins in complex biological samples. In computational p

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Robust Parameter Learning for Uncertain MDPs

DGX agent

arXiv:2605.01339v1 Announce Type: new Abstract: Learning-based approaches to verifying unknown Markov decision processes (MDPs) often employ uncertain MDPs. These models use, for example, confidence i

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

SF20K Competition 2025: Summary and findings

DGX agent

arXiv:2605.01496v1 Announce Type: new Abstract: This report presents the results and findings of the first edition of the Short-Films 20K (SF20K) Competition, held in conjunction with the SLoMO Worksh

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection

DGX agent

arXiv:2605.02860v1 Announce Type: cross Abstract: Cross-language code clone detection (X-CCD) is challenging because semantically equivalent programs written in different languages often share little

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Subquadratic launches with $29M to bring 12M-token context windows to AI

DGX agent

Subquadratic, a company developing a novel generative artificial intelligence model, launched today with 29 million in seed funding. The new large language model, dubbed SubQ, uses what the company ca

model-releasessiliconangle
5 May 2026
Model Releases

The Compliance Trap: How Structural Constraints Degrade Frontier AI Metacognition Under Adversarial Pressure

DGX agent

arXiv:2605.02398v1 Announce Type: cross Abstract: As frontier AI models are deployed in high-stakes decision pipelines, their ability to maintain metacognitive stability -- knowing what they do not kn

model-releasesarxiv-cs-cl
5 May 2026
Safety

The Model Knows, the Decoder Finds: Future Value Guided Particle Power Sampling

DGX agent

arXiv:2605.02427v1 Announce Type: cross Abstract: A recurring pattern in 'reasoning without training' is that base LLMs already assign non-trivial probability mass to correct multi-step solutions; the

safetyarxiv-cs-lg
5 May 2026
Model Releases

Understanding the Performance Plateau in Text-to-Video Retrieval: A Comprehensive Empirical and Linguistic Analysis

DGX agent

arXiv:2605.00826v1 Announce Type: cross Abstract: Text-to-video retrieval enables users to find relevant video content using natural language queries, a task that has grown increasingly important with

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Alethia: A Foundational Encoder for Voice Deepfakes

DGX agent

arXiv:2605.00251v1 Announce Type: cross Abstract: Existing voice deepfake detection and localization models rely heavily on representations extracted from speech foundation models (SFMs). However, dow

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Caracal: Causal Architecture via Spectral Mixing

DGX agent

arXiv:2605.00292v1 Announce Type: new Abstract: The scalability of Large Language Models to long sequences is hindered by the quadratic cost of attention and the limitations of positional encodings. T

model-releasesarxiv-cs-lg
4 May 2026
Research

Embodied Interpretability: Linking Causal Understanding to Generalization in Vision-Language-Action Models

DGX agent

arXiv:2605.00321v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies often fail under distribution shift, suggesting that decisions may depend on spurious visual correlations rather t

researcharxiv-cs-ro
4 May 2026
Model Releases

Evaluating the Architectural Reasoning Capabilities of LLM Provers via the Obfuscated Natural Number Game

DGX agent

arXiv:2605.00677v1 Announce Type: new Abstract: While Large Language Models have achieved notable success on formal mathematics benchmarks such as MiniF2F, it remains unclear whether these results ste

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

From Prediction to Practice: A Task-Aware Evaluation Framework for Blood Glucose Forecasting

DGX agent

arXiv:2605.00645v1 Announce Type: new Abstract: Clinical time-series forecasting is increasingly studied for decision support, yet standard aggregate metrics can obscure whether a model is actually us

model-releasesarxiv-cs-lg
4 May 2026
Research

Generative Modeling under Non-Monotone MAR Missingness via Approximate Wasserstein Gradient Flows

DGX agent

arXiv:2604.04567v2 Announce Type: replace-cross Abstract: The prevalence of missing values in data science poses a substantial risk to any further analyses. Despite a wealth of research, principled no

researcharxiv-cs-lg
4 May 2026
Model Releases

How I built a free, local AI powerhouse in 10 days (Ollama + Gemma 4 + Claude Cowork 3P + Browserless)

DGX agent

This post documents a 10-day project to build a local AI system using open-source tools and models, specifically combining Ollama (a local LLM framework), Gemma 4 (a language model), Claude Cowork 3P,

model-releasesr-ollama
4 May 2026
Safety

Model-Based Reinforcement Learning with Double Oracle Efficiency in Policy Optimization and Offline Estimation

DGX agent

arXiv:2605.00393v1 Announce Type: new Abstract: Reinforcement learning (RL) in large environments often suffers from severe computational bottlenecks, as conventional regret minimization algorithms re

safetyarxiv-cs-lg
4 May 2026
Industry

The foundation of AI scalability: one team, one platform, one operating model

DGX agent

This Databricks blog post discusses how organizations can achieve AI scalability through unified infrastructure and organizational alignment, emphasizing the importance of consolidating teams, platfor

industrydatabricks
4 May 2026
Model Releases

Qwen3.6 vs gpt-oss:120b on Apple Silicon — three Qwen variants benchmarked, plus what works and where it does not

DGX agent

This post benchmarks three Qwen3.6 model variants against gpt-oss:120b when running on Apple Silicon hardware, evaluating their performance characteristics and practical usability. It documents both t

model-releasesr-ollama
3 May 2026
Research

ActiNet: An Open-Source Tool for Activity Intensity Classification of Wrist-Worn Accelerometry Using Self-Supervised Deep Learning

DGX agent

arXiv:2510.01712v2 Announce Type: replace Abstract: The use of accurate and reliable open-source human activity recognition (HAR) models on passively collected wrist-accelerometer data is essential in

researcharxiv-cs-lg
1 May 2026
Model Releases

Beyond Accuracy: LLM Variability in Evidence Screening for Software Engineering SLRs

DGX agent

arXiv:2604.27006v1 Announce Type: cross Abstract: Context: Study screening in systematic literature reviews is costly, inconsistency-prone, and risk-asymmetric, since false negatives can compromise va

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Characterizing the Consistency of the Emergent Misalignment Persona

DGX agent

arXiv:2604.28082v1 Announce Type: new Abstract: Fine-tuning large language models (LLMs) on narrowly misaligned data generalizes to broadly misaligned behavior, a phenomenon termed emergent misalignme

model-releasesarxiv-cs-ai
1 May 2026
Research

DeepWeightFlow: Re-Basined Flow Matching for Generating Neural Network Weights

DGX agent

arXiv:2601.05052v2 Announce Type: replace Abstract: Building efficient and effective generative models for neural network weights has been a research focus of significant interest that faces challenge

researcharxiv-cs-lg
1 May 2026
Safety

Design Structure Matrix Modularization with Large Language Models

DGX agent

arXiv:2604.28018v1 Announce Type: cross Abstract: Design Structure Matrix (DSM) modularization, the task of partitioning system elements into cohesive modules, is a fundamental combinatorial challenge

safetyarxiv-cs-ai
1 May 2026
Model Releases

Do What I Say: A Spoken Prompt Dataset for Instruction-Following

DGX agent

arXiv:2603.09881v2 Announce Type: replace Abstract: Speech Large Language Models (SLLMs) have rapidly expanded, supporting a wide range of tasks. These models are typically evaluated using text prompt

model-releasesarxiv-cs-cl
1 May 2026
Local Ai

Enhancing Linux Privilege Escalation Attack Capabilities of Local LLM Agents

DGX agent

arXiv:2604.27143v1 Announce Type: cross Abstract: Recent research has demonstrated the potential of Large Language Models (LLMs) for autonomous penetration testing, particularly when using cloud-based

local-aiarxiv-cs-ai
1 May 2026
Research

Fitting Horn DL Ontologies to ABox and Query Examples: A Tale of Simulation Quantifiers and Finite Models

DGX agent

arXiv:2604.26976v1 Announce Type: cross Abstract: We study the problem of fitting a description logic (DL) ontology to a given set of positive and negative examples that take the form of an ABox and a

researcharxiv-cs-ai
1 May 2026
Model Releases

MiniCPM-o 4.5: Towards Real-Time Full-Duplex Omni-Modal Interaction

DGX agent

arXiv:2604.27393v1 Announce Type: new Abstract: Recent progress in multimodal large language models (MLLMs) has brought AI capabilities from static offline data processing to real-time streaming inter

model-releasesarxiv-cs-cl
1 May 2026
Research

Modeling Spatial Extremal Dependence of Precipitation Using Distributional Neural Networks

DGX agent

arXiv:2407.08668v3 Announce Type: replace-cross Abstract: In this work, we propose a simulation-based estimation approach using generative neural networks to determine dependencies of precipitation ma

researcharxiv-cs-lg
1 May 2026
Research

Musk v. Altman week 1: Elon Musk says he was duped, warns AI could kill us all, and admits that xAI distills OpenAI’s models

DGX agent

In the first week of the landmark trial between Elon Musk and OpenAI, Musk took the stand in a crisp black suit and tie and argued that OpenAI CEO Sam Altman and president Greg Brockman had deceived h

researchmit-tech-review
1 May 2026
Research

Optimization before Evaluation: Evaluation with Unoptimised Prompts Can be Misleading

DGX agent

arXiv:2604.27637v1 Announce Type: new Abstract: Current Large Language Model (LLM) evaluation frameworks utilize the same static prompt template across all models under evaluation. This differs from t

researcharxiv-cs-ai
1 May 2026
Model Releases

Perturbation Probing: A Two-Pass-per-Prompt Diagnostic for FFN Behavioral Circuits in Aligned LLMs

DGX agent

arXiv:2604.27401v1 Announce Type: new Abstract: Perturbation probing generates task-specific causal hypotheses for FFN neurons in large language models using two forward passes per prompt and no backp

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

PVeRA: Probabilistic Vector-Based Random Matrix Adaptation

DGX agent

arXiv:2512.07703v2 Announce Type: replace Abstract: Large foundation models have emerged in the last years and are pushing performance boundaries for a variety of tasks. Training or even finetuning su

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension

DGX agent

arXiv:2601.14289v2 Announce Type: replace-cross Abstract: Understanding research papers remains challenging for foundation models due to specialized scientific discourse and complex figures and tables

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

RuC: HDL-Agnostic Rule Completion Benchmark Generation

DGX agent

arXiv:2604.27780v1 Announce Type: cross Abstract: Large Language Models (LLMs) have rapidly improved in performance across code-related tasks, making their integration into Register Transfer Level (RT

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

SpecVQA: A Benchmark for Spectral Understanding and Visual Question Answering in Scientific Images

DGX agent

arXiv:2604.28039v1 Announce Type: new Abstract: Spectra are a prevalent yet highly information-dense form of scientific imagery, presenting substantial challenges to multimodal large language models (

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

A Practice of Post-Training on Llama-3 70B with Optimal Selection of Additional Language Mixture Ratio

DGX agent

arXiv:2409.06624v4 Announce Type: replace-cross Abstract: Large Language Models (LLM) often need to be Continual Pre-Trained (CPT) to obtain unfamiliar language skills or adapt to new domains. The hug

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

A Systematic Comparison of Prompting and Multi-Agent Methods for LLM-based Stance Detection

DGX agent

arXiv:2604.26319v1 Announce Type: new Abstract: Stance detection identifies the attitude of a text author toward a given target. Recent studies have explored various LLM-based strategies for this task

model-releasesarxiv-cs-cl
30 Apr 2026
Tutorials

AWS Generative AI Model Agility Solution: A comprehensive guide to migrating LLMs for generative AI production

DGX agent

In this post, we introduce a systematic framework for LLM migration or upgrade in generative AI production, encompassing essential tools, methodologies, and best practices. The framework facilitates t

tutorialsaws-ml-blog
30 Apr 2026
Model Releases

> be me > 'the internet is polluted by ai slop, we need low-background tokens' > 'wouldnt it be cool if we could time travel and see what ou…

DGX agent

> be me > 'the internet is polluted by ai slop, we need low-background tokens' > 'wouldnt it be cool if we could time travel and see what our ancestors 100 years ago would say to us' > all the existin

model-releasesswyx--x
30 Apr 2026
Research

Bootstrapping Sign Language Annotations with Sign Language Models

DGX agent

AI-driven sign language interpretation is limited by a lack of high-quality annotated data. New datasets including ASL STEM Wiki and FLEURS-ASL contain professional interpreters and 100s of hours of d

researchapple-ml-research
30 Apr 2026
Research

Budget-Constrained Causal Bandits: Bridging Uplift Modeling and Sequential Decision-Making

DGX agent

arXiv:2604.26169v1 Announce Type: new Abstract: Treatment allocation under budget constraints is a central challenge in digital advertising: advertisers must decide which users to show ads to while sp

researcharxiv-cs-lg
30 Apr 2026
Model Releases

CoQuant: Joint Weight-Activation Subspace Projection for Mixed-Precision LLMs

DGX agent

arXiv:2604.26378v1 Announce Type: new Abstract: Post-training quantization (PTQ) has become an important technique for reducing the inference cost of Large Language Models (LLMs). While recent mixed-p

model-releasesarxiv-cs-lg
30 Apr 2026
Local Ai

Emergent Coordination in Multi-Agent Language Models

DGX agent

arXiv:2510.05174v4 Announce Type: replace-cross Abstract: When are multi-agent LLM systems merely a collection of individual agents versus an integrated collective with higher-order structure? We intr

local-aiarxiv-cs-ai
30 Apr 2026
Model Releases

MoRFI: Monotonic Sparse Autoencoder Feature Identification

DGX agent

arXiv:2604.26866v1 Announce Type: new Abstract: Large language models (LLMs) acquire most of their factual knowledge during the pre-training stage, through next token prediction. Subsequent stages of

model-releasesarxiv-cs-cl
30 Apr 2026
← Previous
1…368369370371372…1326
Next →