AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,597 results
24 Apr 2026

Prototype-Based Test-Time Adaptation of Vision-Language Models

ResearchDGX agent

arXiv:2604.21360v1 Announce Type: new Abstract: Test-time adaptation (TTA) has emerged as a promising paradigm for vision-language models (VLMs) to bridge the distribution gap between pre-training and

We really need a better word for the good kind of AI psychosis, the one where someone goes into a fugue state with the latest model and retu…

ApplicationsDGX agent

Ethan Mollick discusses the phenomenon of users entering deeply immersive, flow-like states while intensely interacting with advanced AI models, suggesting the need for better terminology to describe

23 Apr 2026

A Computational Model of Message Sensation Value in Short Video Multimodal Features that Predicts Sensory and Behavioral Engagement

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
ResearchDGX agent

arXiv:2604.19995v1 Announce Type: new Abstract: The contemporary media landscape is characterized by sensational short videos. While prior research examines the effects of individual multimodal featur

CyberCertBench: Evaluating LLMs in Cybersecurity Certification Knowledge

Model ReleasesDGX agent

arXiv:2604.20389v1 Announce Type: cross Abstract: The rapid evolution and use of Large Language Models (LLMs) in professional workflows require an evaluation of their domain-specific knowledge against

Deepseek V4 on AI Gateway

Model ReleasesDGX agent

Vercel announced support for DeepSeek V4 model through its AI Gateway service, enabling developers to access this AI model alongside other supported models on the platform. This addition expands Verce

Exploring Data Augmentation and Resampling Strategies for Transformer-Based Models to Address Class Imbalance in AI Scoring of Scientific Explanations in NGSS Classroom

Model ReleasesDGX agent

arXiv:2604.19754v1 Announce Type: new Abstract: Automated scoring of students' scientific explanations offers the potential for immediate, accurate feedback, yet class imbalance in rubric categories p

Kalman Filter Enhanced GRPO for Reinforcement Learning-Based Language Model Reasoning

SafetyDGX agent

arXiv:2505.07527v5 Announce Type: replace Abstract: The advantage function is a central concept in RL that helps reduce variance in policy gradient estimates. For language modeling, Group Relative Pol

Latent Stochastic Interpolants

Model ReleasesDGX agent

arXiv:2506.02276v2 Announce Type: replace Abstract: Stochastic Interpolants (SI) is a powerful framework for generative modeling, capable of flexibly transforming between two probability distributions

Maximum Likelihood Reconstruction for Multi-Look Digital Holography with Markov-Modeled Speckle Correlation

ResearchDGX agent

arXiv:2604.20154v1 Announce Type: cross Abstract: Multi-look acquisition is a widely used strategy for reducing speckle noise in coherent imaging systems such as digital holography. By acquiring multi

Object Referring-Guided Scanpath Prediction with Perception-Enhanced Vision-Language Models

ResearchDGX agent

arXiv:2604.20361v1 Announce Type: new Abstract: Object Referring-guided Scanpath Prediction (ORSP) aims to predict the human attention scanpath when they search for a specific target object in a visua

Representational Alignment Across Model Layers and Brain Regions with Multi-Level Optimal Transport

SafetyDGX agent

arXiv:2510.01706v2 Announce Type: replace-cross Abstract: Standard representational similarity methods align each layer of a network to its best match in another independently, producing asymmetric re

ThermoQA: A Three-Tier Benchmark for Evaluating Thermodynamic Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2604.19758v1 Announce Type: new Abstract: We present ThermoQA, a benchmark of 293 open-ended engineering thermodynamics problems in three tiers: property lookups (110 Q), component analysis (101

Tracing Relational Knowledge Recall in Large Language Models

ResearchDGX agent

arXiv:2604.19934v1 Announce Type: new Abstract: We study how large language models recall relational knowledge during text generation, with a focus on identifying latent representations suitable for r

Variance Is Not Importance: Structural Analysis of Transformer Compressibility Across Model Scales

Model ReleasesDGX agent

arXiv:2604.20682v1 Announce Type: new Abstract: We present a systematic empirical study of transformer compression through over 40 experiments on GPT-2 (124M parameters) and Mistral 7B (7.24B paramete

we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a gre…

HardwareDGX agent

we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a great business strategy 🙈 So now we have 2 free opus sessions/d

22 Apr 2026

ARM: Advantage Reward Modeling for Long-Horizon Manipulation

SafetyDGX agent

arXiv:2604.03037v2 Announce Type: replace-cross Abstract: Long-horizon robotic manipulation remains challenging for reinforcement learning (RL) because sparse rewards provide limited guidance for cred

Assessing Capabilities of Large Language Models in Social Media Analytics: A Multi-task Quest

Model ReleasesDGX agent

arXiv:2604.18955v1 Announce Type: cross Abstract: In this study, we present the first comprehensive evaluation of modern LLMs - including GPT-4, GPT-4o, GPT-3.5-Turbo, Gemini 1.5 Pro, DeepSeek-V3, Lla

Attention-based Multi-modal Deep Learning Model of Spatio-temporal Crop Yield Prediction with Satellite, Soil and Climate Data

SafetyDGX agent

arXiv:2604.19217v1 Announce Type: cross Abstract: Crop yield prediction is one of the most important challenge, which is crucial to world food security and policy-making decisions. The conventional fo

Can an installed local model have access to my pc?

Local AiDGX agent

By default, Ollama binds to port 11434 on localhost, which means it's accessible only on your own machine. Ollama has no built-in authentication and should be secured with firewall rules, VPN, or reve

Decomposed Trust: Privacy, Adversarial Robustness, Ethics, and Fairness in Low-Rank LLMs

SafetyDGX agent

arXiv:2511.22099v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have driven major advances across domains, yet their massive size hinders deployment in resource-constrained sett

Enabling the agentic enterprise: business and industry agents arrive in Gemini Enterprise

Model ReleasesDGX agent

We are officially moving past the era of one-size-fits-all AI. Enterprises today require highly specialized, role-specific tools to drive real productivity — but they cannot afford to sacrifice securi

Heterogeneity-Aware Personalized Federated Learning for Industrial Predictive Analytics

Model ReleasesDGX agent

arXiv:2604.19451v1 Announce Type: new Abstract: Federated prognostics enable clients (e.g., companies, factories, and production lines) to collaboratively develop a failure time prediction model while

Introducing Kimi K2.6 from @Kimi_Moonshot, a multimodal agentic model with Agent Swarm scaling to 300 sub-agents and long-horizon coding sta…

AgentsDGX agent

Introducing Kimi K2.6 from @Kimi_Moonshot, a multimodal agentic model with Agent Swarm scaling to 300 sub-agents and long-horizon coding stability. AI natives can now use Kimi K2.6 on Together AI and

Location Not Found: Exposing Implicit Local and Global Biases in Multilingual LLMs

SafetyDGX agent

arXiv:2604.19292v1 Announce Type: cross Abstract: Multilingual large language models (LLMs) have minimized the fluency gap between languages. This advancement, however, exposes models to the risk of b

Lost in Translation: Do LVLM Judges Generalize Across Languages?

Model ReleasesDGX agent

arXiv:2604.19405v1 Announce Type: new Abstract: Automatic evaluators such as reward models play a central role in the alignment and evaluation of large vision-language models (LVLMs). Despite their gr

Sadly, this post will result in future AI models being “squashmaxxed” - good at producing butternut squash images, bad at everything else.

ApplicationsDGX agent

This post discusses concerns about AI model degradation through the inclusion of low-quality or irrelevant training data, using butternut squash image generation as a humorous example of how excessive

SAGE-32B: Agentic Reasoning via Iterative Distillation

Model ReleasesDGX agent

arXiv:2601.04237v2 Announce Type: replace Abstract: We demonstrate SAGE-32B, a 32 billion parameter language model that focuses on agentic reasoning and long range planning tasks. Unlike chat models t

The High Explosives and Affected Targets (HEAT) Dataset

Model ReleasesDGX agent

arXiv:2604.18828v1 Announce Type: new Abstract: Artificial Intelligence (AI) surrogate models provide a computationally efficient alternative to full-physics simulations, but no public datasets curren

This pipeline is why the same base model produces more accurate, better-cited, and more efficient answers inside Perplexity than out of the …

ToolsDGX agent

This pipeline is why the same base model produces more accurate, better-cited, and more efficient answers inside Perplexity than out of the box. Read our research: https://research.perplexity.ai/artic

What’s new in BigQuery: Powering the Agentic Era

Model ReleasesDGX agent

Succeeding in the agentic era requires a transformation in your data strategy: moving from human-scale to agent-first workloads, evolving from reactive intelligence to proactive action, and shifting f

21 Apr 2026

A Survey of Reinforcement Learning for Large Language Models under Data Scarcity: Challenges and Solutions

TutorialsDGX agent

arXiv:2604.17312v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a powerful post-training paradigm for enhancing the reasoning capabilities of large language models (LLMs). H

AI Insider: 'Adding a Human Makes Your Team Worse' Emad Mostaque | @EMostaque TIMESTAMPS : 00:00 The models too dangerous to release 10:06 W…

IndustryDGX agent

AI Insider: 'Adding a Human Makes Your Team Worse' Emad Mostaque | @EMostaque TIMESTAMPS : 00:00 The models too dangerous to release 10:06 Why physics needs axioms — and AI doesn't 11:43 The MIND fram

AVRT: Audio-Visual Reasoning Transfer through Single-Modality Teachers

Model ReleasesDGX agent

arXiv:2604.16617v1 Announce Type: new Abstract: Recent advances in reasoning models have shown remarkable progress in text-based domains, but transferring those capabilities to multimodal settings, e.

Bielik Guard: Efficient Polish Language Safety Classifiers for LLM Content Moderation

Model ReleasesDGX agent

arXiv:2602.07954v4 Announce Type: replace Abstract: As Large Language Models (LLMs) become increasingly deployed in Polish language applications, the need for efficient and accurate content safety cla

CLAG: Adaptive Memory Organization via Agent-Driven Clustering for Small Language Model Agents

AgentsDGX agent

arXiv:2603.15421v2 Announce Type: replace Abstract: Large language model agents heavily rely on external memory to support knowledge reuse and complex reasoning tasks. Yet most memory systems store ex

Cognitive Chain-of-Thought (CoCoT): Structured Multimodal Reasoning about Social Situations

Model ReleasesDGX agent

arXiv:2507.20409v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting helps models think step by step. But naive CoT breaks down in visually grounded social tasks, where models must per

Compositional Steering of Large Language Models with Steering Tokens

ApplicationsDGX agent

arXiv:2601.05062v2 Announce Type: replace Abstract: Deploying LLMs in real-world applications requires controllable output that satisfies multiple desiderata at the same time. While existing work exte

D-Prism: Differentiable Primitives for Structured Dynamic Modeling

ResearchDGX agent

arXiv:2604.17082v1 Announce Type: new Abstract: Capturing both geometry and rigid motion for structured dynamic objects, like multi-part assemblies or jointed mechanisms, remains a key challenge. Exis

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models

SafetyDGX agent

arXiv:2511.15669v2 Announce Type: replace Abstract: Does Chain-of-Thought (CoT) reasoning genuinely improve Vision-Language-Action (VLA) models, or does it merely add overhead? Existing CoT-VLA system

Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion

ResearchDGX agent

arXiv:2604.16656v1 Announce Type: new Abstract: All languages are equal; when it comes to tokenization, some are more equal than others. Tokens are the hidden currency that dictate the cost and latenc

emg2speech: Synthesizing speech from electromyography using self-supervised speech models

ResearchDGX agent

arXiv:2510.23969v2 Announce Type: replace-cross Abstract: We present a neuromuscular speech interface that translates electromyographic (EMG) signals recorded from orofacial muscles during speech arti

From Clinical Intent to Clinical Model: An Autonomous Coding-Agent Framework for Clinician-driven AI Development

AgentsDGX agent

arXiv:2604.17110v1 Announce Type: new Abstract: Clinical AI development has traditionally followed a collaborative paradigm that depends on close interaction between clinicians and specialized AI team

From Handwriting to Structured Data: Benchmarking AI Digitisation of Handwritten Forms

Model ReleasesDGX agent

arXiv:2604.16504v1 Announce Type: new Abstract: Manual digitisation of structured handwritten documents is slow and costly. We benchmark 17 leading frontier multi-modal large language models and open-

How Robustly do LLMs Understand Execution Semantics?

Model ReleasesDGX agent

arXiv:2604.16320v1 Announce Type: cross Abstract: LLMs demonstrate remarkable reasoning capabilities, yet whether they utilize internal world models or rely on sophisticated pattern matching remains o

it's entirely fine and great for proprietary models to exist the issue is they try to kill open source behind scenes while pretending they s…

IndustryDGX agent

it's entirely fine and great for proprietary models to exist the issue is they try to kill open source behind scenes while pretending they support it so far they've been pretty incompetent but they'll

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is…

HardwareDGX agent

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is just scratching the surface. Ultimately, it boils down to g

Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users

Model ReleasesDGX agent

arXiv:2603.16120v2 Announce Type: replace Abstract: Deep Research (DR) systems help researchers cope with ballooning publishing counts. Such tools synthesize scientific papers to answer research queri

Latent Preference Modeling for Cross-Session Personalized Tool Calling

Model ReleasesDGX agent

arXiv:2604.17886v1 Announce Type: new Abstract: Users often omit essential details in their requests to LLM-based agents, resulting in under-specified inputs for tool use. This poses a fundamental cha

Long-Text-to-Image Generation via Compositional Prompt Decomposition

Model ReleasesDGX agent

arXiv:2604.18258v1 Announce Type: new Abstract: While modern text-to-image (T2I) models excel at generating images from intricate prompts, they struggle to capture the key details when the inputs are

LoReC: Rethinking Large Language Models for Graph Data Analysis

ResearchDGX agent

arXiv:2604.17897v1 Announce Type: new Abstract: The advent of Large Language Models (LLMs) has fundamentally reshaped the way we interact with graphs, giving rise to a new paradigm called GraphLLM. As

Modeling, Control and Self-sensing of Dielectric Elastomer Soft Actuators: A Review

ResearchDGX agent

arXiv:2604.17199v1 Announce Type: new Abstract: Dielectric elastomer actuators (DEAs) have garnered extensive attention especially in soft robotic applications over the past few decades owing to the a

Modeling Higher-Order Brain Interactions via a Multi-View Information Bottleneck Framework for fMRI-based Psychiatric Diagnosis

Model ReleasesDGX agent

arXiv:2604.17713v1 Announce Type: new Abstract: Resting-state functional magnetic resonance imaging (fMRI) has emerged as a cornerstone for psychiatric diagnosis, yet most approaches rely on pairwise

Modelling Gas-Phase Reaction Kinetics with Guided Particle Diffusion Sampling

Model ReleasesDGX agent

arXiv:2604.16461v1 Announce Type: cross Abstract: Physics-guided sampling with diffusion priors has recently shown strong performance in solving complex systems of partial differential equations (PDEs

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

SafetyDGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

PiCa: Parameter-Efficient Fine-Tuning with Column Space Projection

Model ReleasesDGX agent

arXiv:2505.20211v3 Announce Type: replace Abstract: Fine-tuning large foundation models is essential for building expert models tailored to specialized tasks and domains, but fully updating billions o

Safer Trajectory Planning with CBF-guided Diffusion Model for Unmanned Aerial Vehicles

SafetyDGX agent

arXiv:2604.17527v1 Announce Type: new Abstract: Safe and agile trajectory planning is essential for autonomous systems, especially during complex aerobatic maneuvers. Motivated by the recent success o

Shifting the Gradient: Understanding How Defensive Training Methods Protect Language Model Integrity

ApplicationsDGX agent

arXiv:2604.16423v1 Announce Type: new Abstract: Defensive training methods such as positive preventative steering (PPS) and inoculation prompting (IP) offer surprising results through seemingly simila

SinkRouter: Sink-Aware Routing for Efficient Long-Context Decoding in Large Language and Multimodal Models

Model ReleasesDGX agent

arXiv:2604.16883v1 Announce Type: new Abstract: In long-context decoding for LLMs and LMMs, attention becomes increasingly memory-bound because each decoding step must load a large amount of KV-cache

Spotlights and Blindspots: Evaluation Machine-Generated Text Detection

ResearchDGX agent

arXiv:2604.16607v1 Announce Type: new Abstract: With the rise of generative language models, machine-generated text detection has become a critical challenge. A wide variety of models is available, bu

Target Parameterization in Diffusion Models for Nonlinear Spatiotemporal System Identification

ResearchDGX agent

arXiv:2604.17566v1 Announce Type: cross Abstract: Machine learning is becoming increasingly important for nonlinear system identification, including dynamical systems with spatially distributed output

← Previous
1…191192193194195…1010
Next →