AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,575 results
21 Apr 2026

Jailbreaking Large Language Models with Morality Attacks

SafetyDGX agent

arXiv:2604.17053v1 Announce Type: new Abstract: Pluralism alignment with AI has the sophisticated and necessary goal of creating AI that can coexist with and serve morally multifaceted humanity. Resea

Latent-Compressed Variational Autoencoder for Video Diffusion Models

ResearchDGX agent

arXiv:2604.16479v1 Announce Type: new Abstract: Video variational autoencoders (VAEs) used in latent diffusion models typically require a sufficiently large number of latent channels to ensure high-qu

Linking Exteroception and Proprioception through Improved Contact Modeling for Soft Growing Robots

ResearchDGX agent

arXiv:2507.10694v2 Announce Type: replace Abstract: Passive deformation due to compliance is a commonly used benefit of soft robots, providing opportunities to achieve robust actuation with few active

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Method for Aggregating Unstructured Data Using Large Language Models

Model ReleasesDGX agent

arXiv:2604.16425v1 Announce Type: cross Abstract: This paper presents a method for the automated collection and aggregation of unstructured data from diverse web sources, utilizing Large Language Mode

Model in Distress: Sentiment Analysis on French Synthetic Social Media

Model ReleasesDGX agent

arXiv:2604.18226v1 Announce Type: new Abstract: Automated analysis of customer feedback on social media is hindered by three challenges: the high cost of annotated training data, the scarcity of evalu

Predictive Modeling of Natural Medicinal Compounds for Alzheimer Disease Using Cheminformatics

ResearchDGX agent

arXiv:2604.18316v1 Announce Type: cross Abstract: The most common cause of dementia is Alzheimer disease, a progressive neurodegenerative disorder affecting older adults that gradually impairs memory,

Reasoning on the Manifold: Bidirectional Consistency for Self-Verification in Diffusion Language Models

SafetyDGX agent

arXiv:2604.16565v1 Announce Type: new Abstract: While Diffusion Large Language Models (dLLMs) offer structural advantages for global planning, efficiently verifying that they arrive at correct answers

Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contrastive Scoring

SafetyDGX agent

arXiv:2512.12069v3 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) are vulnerable to a growing array of multimodal jailbreak attacks, necessitating defenses that are both g

SERM: Self-Evolving Relevance Model with Agent-Driven Learning from Massive Query Streams

AgentsDGX agent

arXiv:2601.09515v2 Announce Type: replace Abstract: Due to the dynamically evolving nature of real-world query streams, relevance models struggle to generalize to practical search scenarios. A sophist

Source: a handful of unauthorized users in a private Discord channel have been accessing Anthropic's Mythos model since the day the company announced it (Rachel Metz/Bloomberg)

IndustryDGX agent

Rachel Metz / Bloomberg: Source: a handful of unauthorized users in a private Discord channel have been accessing Anthropic's Mythos model since the day the company announced it — A small group of una

SPaRSe-TIME: Saliency-Projected Low-Rank Temporal Modeling for Efficient and Interpretable Time Series Prediction

ApplicationsDGX agent

arXiv:2604.17350v1 Announce Type: cross Abstract: Time series forecasting is traditionally dominated by sequence-based architectures such as recurrent neural networks and attention mechanisms, which p

SpidR-Adapt: A Universal Speech Representation Model for Few-Shot Adaptation

ResearchDGX agent

arXiv:2512.21204v2 Announce Type: replace Abstract: Human infants, with only a few hundred hours of speech exposure, acquire basic units of new languages, highlighting a striking efficiency gap compar

The Gait Signature of Frailty: Transfer Learning based Deep Gait Models for Scalable Frailty Assessment

ResearchDGX agent

arXiv:2603.24434v2 Announce Type: replace Abstract: Frailty is a condition in aging medicine characterized by diminished physiological reserve and increased vulnerability to stressors. However, frailt

TokenChain: A Discrete Speech Chain via Semantic Token Modeling

ApplicationsDGX agent

arXiv:2510.06201v1 Announce Type: cross Abstract: Machine Speech Chain, simulating the human perception-production loop, proves effective in jointly improving ASR and TTS. We propose TokenChain, a ful

VocabTailor: Dynamic Vocabulary Selection for Downstream Tasks in Small Language Models

Local AiDGX agent

arXiv:2508.15229v3 Announce Type: replace Abstract: Small Language Models (SLMs) provide computational advantages in resource-constrained environments, yet memory limitations remain a critical bottlen

Where's the raccoon with the ham radio? (ChatGPT Images 2.0)

Model ReleasesDGX agent

OpenAI released ChatGPT Images 2.0 today, their latest image generation model. On the livestream Sam Altman said that the leap from gpt-image-1 to gpt-image-2 was equivalent to jumping from GPT-3 to G

20 Apr 2026

A useful ward against slop story/science posts on X is noting which is in the character limit. All of the models struggle to do 280 characte…

ApplicationsDGX agent

A useful ward against slop story/science posts on X is noting which is in the character limit. All of the models struggle to do 280 character summaries on their first pass, and most of the people crea

Anthropomorphism and Trust in Human-Large Language Model interactions

ResearchDGX agent

arXiv:2604.15316v1 Announce Type: cross Abstract: With large language models (LLMs) becoming increasingly prevalent in daily life, so too has the tendency to attribute to them human-like minds and emo

Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning

Local AiDGX agent

arXiv:2604.15414v1 Announce Type: cross Abstract: Continual reinforcement learning must balance retention with adaptation, yet many methods still rely on single-model preservation, committing to one e

CiPO: Counterfactual Unlearning for Large Reasoning Models through Iterative Preference Optimization

ResearchDGX agent

arXiv:2604.15847v1 Announce Type: new Abstract: Machine unlearning has gained increasing attention in recent years, as a promising technique to selectively remove unwanted privacy or copyrighted infor

Convolutionally Low-Rank Models with Modified Quantile Regression for Interval Time Series Forecasting

ApplicationsDGX agent

arXiv:2604.15791v1 Announce Type: new Abstract: The quantification of uncertainty in prediction models is crucial for reliable decision-making, yet remains a significant challenge. Interval time serie

Dataset 2: Drift, named after Drift Diffusion Modeling. https://huggingface.co/datasets/DJLougen/drift-preview-5k This is a bit of different…

IndustryDGX agent

Dataset 2 (Drift) is a 5,000-sample preview dataset hosted on Hugging Face, named after drift diffusion modeling concepts. The dataset appears to represent a novel or experimental approach to data col

Information-Consistent Language Model Recommendations through Group Relative Policy Optimization

SafetyDGX agent

arXiv:2512.12858v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in business-critical domains such as finance, education, healthcare, and customer suppo

Online Distributionally Robust LLM Alignment via Regression to Relative Reward

Model ReleasesDGX agent

arXiv:2509.19104v2 Announce Type: replace Abstract: Reinforcement Learning with Human Feedback (RLHF) has become crucial for aligning Large Language Models (LLMs) with human intent. However, existing

SwanNLP at SemEval-2026 Task 5: An LLM-based Framework for Plausibility Scoring in Narrative Word Sense Disambiguation

Model ReleasesDGX agent

arXiv:2604.16262v1 Announce Type: new Abstract: Recent advances in language models have substantially improved Natural Language Understanding (NLU). Although widely used benchmarks suggest that Large

19 Apr 2026

Give your local Ollama models a personal knowledge bank (graph-based, not just vector search)

Local AiDGX agent

This post discusses Graph RAG, an approach that uses local LLMs with Ollama to build graph-based knowledge indexes from source documents by deriving entity knowledge graphs and pregenerating community

17 Apr 2026

Acceptance Dynamics Across Cognitive Domains in Speculative Decoding

Model ReleasesDGX agent

arXiv:2604.14682v1 Announce Type: cross Abstract: Speculative decoding accelerates large language model (LLM) inference. It uses a small draft model to propose a tree of future tokens. A larger target

AD4AD: Benchmarking Visual Anomaly Detection Models for Safer Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.15291v1 Announce Type: new Abstract: The reliability of a machine vision system for autonomous driving depends heavily on its training data distribution. When a vehicle encounters significa

Assessing the Potential of Masked Autoencoder Foundation Models in Predicting Downhole Metrics from Surface Drilling Data

ResearchDGX agent

arXiv:2604.15169v1 Announce Type: new Abstract: Oil and gas drilling operations generate extensive time-series data from surface sensors, yet accurate real-time prediction of critical downhole metrics

CURA: Clinical Uncertainty Risk Alignment for Language Model-Based Risk Prediction

SafetyDGX agent

arXiv:2604.14651v1 Announce Type: new Abstract: Clinical language models (LMs) are increasingly applied to support clinical risk prediction from free-text notes, yet their uncertainty estimates often

DA-Cramming: Enhancing Cost-Effective Language Model Pretraining with Dependency Agreement Integration

HardwareDGX agent

arXiv:2311.04799v2 Announce Type: replace Abstract: Pretraining language models is still a challenge for many researchers due to its substantial computational costs. As such, there is growing interest

HAMSA: Scanning-Free Vision State Space Models via SpectralPulseNet

ResearchDGX agent

arXiv:2604.14724v1 Announce Type: new Abstract: Vision State Space Models (SSMs) like Vim, VMamba, and SiMBA rely on complex scanning strategies to adapt sequential SSMs to process 2D images, introduc

LexGenius: An Expert-Level Benchmark for Large Language Models in Legal General Intelligence

Model ReleasesDGX agent

arXiv:2512.04578v3 Announce Type: replace Abstract: Legal general intelligence (GI) refers to artificial intelligence (AI) that encompasses legal understanding, reasoning, and decision-making, simulat

MEBench: A Novel Benchmark for Understanding Mutual Exclusivity Bias in Vision-Language Models

Model ReleasesDGX agent

arXiv:2505.20122v2 Announce Type: replace Abstract: This paper introduces MEBench, a novel benchmark for evaluating mutual exclusivity (ME) bias, a cognitive phenomenon observed in children during wor

Model-Based Reinforcement Learning Exploits Passive Body Dynamics for High-Performance Biped Robot Locomotion

ResearchDGX agent

arXiv:2604.14565v1 Announce Type: new Abstract: Embodiment is a significant keyword in recent machine learning fields. This study focused on the passive nature of the body of a biped robot to generate

Predicting Post-Traumatic Epilepsy from Clinical Records using Large Language Model Embeddings

ResearchDGX agent

arXiv:2604.14547v1 Announce Type: new Abstract: Objective: Post-traumatic epilepsy (PTE) is a debilitating neurological disorder that develops after traumatic brain injury (TBI). Early prediction of P

Reference-Free Sampling-Based Model Predictive Control

HardwareDGX agent

arXiv:2511.19204v3 Announce Type: replace Abstract: We present a sampling-based model predictive control (MPC) framework that enables emergent locomotion without relying on handcrafted gait patterns o

Schema Key Wording as an Instruction Channel in Structured Generation under Constrained Decoding

Model ReleasesDGX agent

arXiv:2604.14862v1 Announce Type: new Abstract: Constrained decoding has been widely adopted for structured generation with large language models (LLMs), ensuring that outputs satisfy predefined forma

The king reigns supreme This is likely going to be the best finetune of qwen 3.6 35b https://huggingface.co/DJLougen/Ornstein3.6-35B-A3B-GGU…

Model ReleasesDGX agent

This post announces Ornstein3.6-35B-A3B-GGU, a fine-tuned variant of Qwen 3.6 35B model available on Hugging Face, with the author claiming it represents a high-quality optimization of the base model.

Variance Computation for Weighted Model Counting with Knowledge Compilation Approach

ApplicationsDGX agent

arXiv:2601.03523v2 Announce Type: replace Abstract: One of the most important queries in knowledge compilation is weighted model counting (WMC), which has been applied to probabilistic inference on va

When Does Content-Based Routing Work? Representation Requirements for Selective Attention in Hybrid Sequence Models

ResearchDGX agent

arXiv:2603.20997v2 Announce Type: replace Abstract: We identify a routing paradox in hybrid sequence models: content-based routing - deciding which tokens deserve expensive attention - requires pairwi

16 Apr 2026

A ghost mechanism: An analytical model of abrupt learning in recurrent networks

Model ReleasesDGX agent

arXiv:2501.02378v2 Announce Type: replace Abstract: Abrupt learning is a common phenomenon in recurrent neural networks (RNNs) trained on working memory tasks. In such cases, the networks develop tran

Adaptive Conformal Prediction for Improving Factuality of Generations by Large Language Models

ResearchDGX agent

arXiv:2604.13991v1 Announce Type: new Abstract: Large language models (LLMs) are prone to generating factually incorrect outputs. Recent work has applied conformal prediction to provide uncertainty es

An Empirical Investigation of Practical LLM-as-a-Judge Improvement Techniques on RewardBench 2

Model ReleasesDGX agent

arXiv:2604.13717v1 Announce Type: new Abstract: LLM-as-a-judge, using a language model to score or rank candidate responses, is widely used as a scalable alternative to human evaluation in RLHF pipeli

Chain of Uncertain Rewards with Large Language Models for Reinforcement Learning

Model ReleasesDGX agent

arXiv:2604.13504v1 Announce Type: cross Abstract: Designing effective reward functions is a cornerstone of reinforcement learning (RL), yet it remains a challenging and labor-intensive process due to

Echoes Over Time: Unlocking Length Generalization in Video-to-Audio Generation Models

SafetyDGX agent

arXiv:2602.20981v3 Announce Type: replace Abstract: Scaling multimodal alignment between video and audio is challenging, particularly due to limited data and the mismatch between text descriptions and

Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchy

Model ReleasesDGX agent

arXiv:2604.02709v2 Announce Type: replace Abstract: The formal reasoning capabilities of LLMs are crucial for advancing automated software engineering. However, existing benchmarks for LLMs lack syste

Flow-based Generative Modeling of Potential Outcomes and Counterfactuals

Model ReleasesDGX agent

arXiv:2505.16051v4 Announce Type: replace-cross Abstract: Predicting potential and counterfactual outcomes from observational data is central to individualized decision-making, particularly in clinica

GPT-Rosalind, our Life Sciences model series, is optimized for scientific workflows, with stronger performance in protein and chemical reaso…

AgentsDGX agent

GPT-Rosalind, our Life Sciences model series, is optimized for scientific workflows, with stronger performance in protein and chemical reasoning, genomics analysis, biochemistry knowledge, and scienti

Stein Variational Uncertainty-Adaptive Model Predictive Control

Model ReleasesDGX agent

arXiv:2604.01034v2 Announce Type: replace Abstract: We propose a Stein variational distributionally robust controller for nonlinear dynamical systems with latent parametric uncertainty. The method is

Text-Attributed Knowledge Graph Enrichment with Large Language Models for Medical Concept Representation

Model ReleasesDGX agent

arXiv:2604.13331v1 Announce Type: new Abstract: In electronic health record (EHR) mining, learning high-quality representations of medical concepts (e.g., standardized diagnosis, medication, and proce

Which video model currently has the best face likeness for LoRA training?

Local AiDGX agent

This r/StableDiffusion thread discusses community comparisons of video generation models (such as those built on Stable Diffusion or FLUX architectures) evaluated specifically for how well they preser

15 Apr 2026

AutoSurrogate: An LLM-Driven Multi-Agent Framework for Autonomous Construction of Deep Learning Surrogate Models in Subsurface Flow

AgentsDGX agent

arXiv:2604.11945v1 Announce Type: cross Abstract: High-fidelity numerical simulation of subsurface flow is computationally intensive, especially for many-query tasks such as uncertainty quantification

[b]=[d]-[t]+[p]: Self-supervised Speech Models Discover Phonological Vector Arithmetic

ResearchDGX agent

arXiv:2602.18899v3 Announce Type: replace-cross Abstract: Self-supervised speech models (S3Ms) are known to encode rich phonetic information, yet how this information is structured remains underexplor

Challenging Vision-Language Models with Physically Deployable Multimodal Semantic Lighting Attacks

SafetyDGX agent

arXiv:2604.12833v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown remarkable performance, yet their security remains insufficiently understood. Existing adversarial studies focu

Characterizing higher-order representations through generative diffusion models explains human decoded neurofeedback performance

TutorialsDGX agent

arXiv:2503.14333v4 Announce Type: replace-cross Abstract: Brains construct not only 'first-order' representations of the environment but also 'higher-order' representations about those representations

Claude Opus 4.7 on Vertex AI

Model ReleasesDGX agent

Today, we’re announcing the general availability of Claude Opus 4.7 on Vertex AI. What’s new: Anthropic’s newest Opus model delivers advanced performance across coding, long-running agents, and profes

ERNIE-Image is now in ComfyUI An open-source 8B DiT text-to-image model from @ErnieforDevs, licensed under Apache-2.0. Key highlights: - Ope…

Local AiDGX agent

ERNIE-Image is now in ComfyUI An open-source 8B DiT text-to-image model from @ErnieforDevs, licensed under Apache-2.0. Key highlights: - Open-source under Apache-2.0 license - Precise multilingual tex

Evaluating Robustness of Large Language Models Against Multilingual Typographical Errors

ApplicationsDGX agent

arXiv:2510.09536v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in multilingual, real-world applications with user inputs -- naturally introducing typographi

Growing Pains: Extensible and Efficient LLM Benchmarking Via Fixed Parameter Calibration

Model ReleasesDGX agent

arXiv:2604.12843v1 Announce Type: new Abstract: The rapid release of both language models and benchmarks makes it increasingly costly to evaluate every model on every dataset. In practice, models are

← Previous
1…165166167168169…1010
Next →