AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Model Releases

Efficient Reasoning with Hidden Thinking

DGX agent

arXiv:2501.19201v2 Announce Type: replace Abstract: Chain-of-Thought (CoT) reasoning has become a powerful framework for improving complex problem-solving capabilities in Multimodal Large Language Mod

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

EmoMM: Benchmarking and Steering MLLM for Multimodal Emotion Recognition under Conflict and Missingness

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.01024v1 Announce Type: new Abstract: Multimodal Emotion Recognition (MER) is critical for interpreting real-world interactions. While Multimodal Large Language Models (MLLM) have shown prom

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Evaluating Tabular Representation Learning for Network Intrusion Detection

DGX agent

arXiv:2605.02519v1 Announce Type: new Abstract: Classic Network Intrusion Detection Systems (NIDS) often rely on manual feature engineering to extract meaningful patterns from network traffic data. Ho

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

FT-RAG: A Fine-grained Retrieval-Augmented Generation Framework for Complex Table Reasoning

DGX agent

arXiv:2605.01495v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) enhances Large Language Models (LLMs) by grounding responses in external knowledge during inference. However, conve

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

GD-FPS: Growth-Driven Feedforward Parameter Selection for Efficient Fine-Tuning

DGX agent

arXiv:2510.27359v2 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) has emerged as a key strategy for adapting large-scale pre-trained models to downstream tasks, but existing a

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

GeoContra: From Fluent GIS Code to Verifiable Spatial Analysis with Geography-Grounded Repair

DGX agent

arXiv:2605.00782v1 Announce Type: cross Abstract: Reliable spatial analysis in GIScience requires preserving coordinate semantics, topology, units, and geographic plausibility. Current LLM-based GIS s

model-releasesarxiv-cs-ai
5 May 2026
Research

Global-Local Feature Decoding with Adapter-Guided SAMv2 for Salient Object Detection

DGX agent

arXiv:2605.02616v1 Announce Type: new Abstract: Salient Object Detection (SOD) remains an essential yet underexplored task in the era of large-scale vision models. Although foundation models like SAM

researcharxiv-cs-cv
5 May 2026
Model Releases

How Well Can We Decode Vowels from Auditory EEG -- A Rigorous Cross-Subject Benchmark with Honest Assessment

DGX agent

arXiv:2605.00865v1 Announce Type: cross Abstract: EEG based phoneme decoding is promising for brain computer interfaces, but many prior studies rely on within subject evaluation, small cohorts, or wea

model-releasesarxiv-cs-cl
5 May 2026
Safety

Implicature in Interaction: Understanding Implicature Improves Alignment in Human-LLM Interaction

DGX agent

arXiv:2510.25426v2 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) is positioning language at the core of human-computer interaction (HCI). We argue that advanci

safetyarxiv-cs-cl
5 May 2026
Model Releases

Implicit Neural Representation-Based Continuous Single Image Super-Resolution: An Empirical Benchmark

DGX agent

arXiv:2601.17723v2 Announce Type: replace Abstract: Implicit neural representation (INR) has become the standard approach for arbitrary-scale image super-resolution (ASSR). To date, no empirical study

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Introducing Agent Gateway ISV ecosystem for security and governance

DGX agent

Managing agents and their actions can quickly grow in complexity and introduce security risks unique to AI. To address these challenges, at Google Cloud Next we announced Agent Gateway to provide simp

model-releasesgoogle-cloud-ai
5 May 2026
Model Releases

Last Week in AI #340 - OpenAI vs Musk + Microsoft, DeepSeek v4, Vision Banana

DGX agent

This newsletter episode covers recent AI industry developments including a legal dispute between OpenAI and Elon Musk, Microsoft's involvement in AI developments, the release of DeepSeek's v4 model, a

model-releaseslast-week-in-ai
5 May 2026
Tutorials

Learning Equivariant Neural-Augmented Object Dynamics From Few Interactions

DGX agent

arXiv:2605.02699v1 Announce Type: cross Abstract: Learning data-efficient object dynamics models for robotic manipulation remains challenging, especially for deformable objects. A popular approach is

tutorialsarxiv-cs-cv
5 May 2026
Research

LinMU: Multimodal Understanding Made Linear

DGX agent

arXiv:2601.01322v2 Announce Type: replace Abstract: Modern Vision-Language Models (VLMs) achieve impressive performance but are limited by the quadratic complexity of self-attention, which prevents th

researcharxiv-cs-cv
5 May 2026
Model Releases

LittleBit-2: Maximizing the Spectral Energy Gain in Sub-1-Bit LLMs via Latent Geometry Alignment

DGX agent

arXiv:2603.00042v2 Announce Type: replace Abstract: We identify the Spectral Energy Gain in extreme model compression, where low-rank binary approximations outperform tiny-rank floating-point baseline

model-releasesarxiv-cs-lg
5 May 2026
Agents

LLM Ghostbusters: Surgical Hallucination Suppression via Adaptive Unlearning

DGX agent

arXiv:2605.01047v1 Announce Type: cross Abstract: Hallucinations, outputs that sound plausible but are factually incorrect, remain an open challenge for deployed LLMs. In code generation, models frequ

agentsarxiv-cs-cl
5 May 2026
Safety

Mitigating Misalignment Contagion by Steering with Implicit Traits

DGX agent

arXiv:2605.02751v1 Announce Type: cross Abstract: Language models (LMs) are increasingly used in high-stakes, multi-agent settings, where following instructions and maintaining value alignment are cri

safetyarxiv-cs-cl
5 May 2026
Research

Mitigating Multimodal LLMs Hallucinations via Relevance Propagation at Inference Time

DGX agent

arXiv:2605.01766v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have revolutionized the landscape of AI, demonstrating impressive capabilities in tackling complex vision and

researcharxiv-cs-cv
5 May 2026
Model Releases

MOSAIC: Multi-agent Orchestration for Task-Intelligent Scientific Coding

DGX agent

arXiv:2510.08804v3 Announce Type: replace Abstract: We present MOSAIC, a multi-agent Large Language Model (LLM) framework for solving challenging scientific coding tasks. Unlike general-purpose coding

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety

DGX agent

arXiv:2605.01687v1 Announce Type: new Abstract: We present MultiBreak, a scalable and diverse multi-turn jailbreak benchmark to evaluate large language model (LLM) safety. Multi-turn jailbreaks mimic

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Pandora's Regret: A Proper Scoring Rule for Evaluating Sequential Search

DGX agent

arXiv:2605.01936v1 Announce Type: new Abstract: In sequential search, alternatives are tested until the true class is found. Standard proper scoring rules like log loss are local, ignoring the ranking

model-releasesarxiv-cs-lg
5 May 2026
Research

ParaRNN: An Interpretable and Parallelizable Recurrent Neural Network for Time-Dependent Data

DGX agent

arXiv:2605.02692v1 Announce Type: cross Abstract: The proliferation of large-scale and structurally complex data has spurred the integration of machine learning methods into statistical modeling. Recu

researcharxiv-cs-lg
5 May 2026
Model Releases

Physics-Informed Neural Learning for State Reconstruction and Parameter Identification in Coupled Greenhouse Climate Dynamics

DGX agent

arXiv:2605.02524v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) have recently emerged as a promising framework for integrating data-driven learning with physical knowledge. In

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Prescriptive Scaling Laws for Data Constrained Training

DGX agent

arXiv:2605.01640v1 Announce Type: cross Abstract: Training compute is increasingly outpacing the availability of high-quality data. This shifts the central challenge from optimal compute allocation to

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Probe-Geometry Alignment: Erasing the Cross-Sequence Memorization Signature Below Chance

DGX agent

arXiv:2605.01699v1 Announce Type: new Abstract: Recent attacks show that behavioural unlearning of large language models leaves internal traces recoverable by adversarial probes. We characterise where

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Projection-Free Transformers via Gaussian Kernel Attention

DGX agent

arXiv:2605.02144v1 Announce Type: new Abstract: Self-attention in Transformers is typically implemented as softmax(QK^op/sqrt{d})V, where Q=XW_Q, K=XW_K, and V=XW_V are learned linear projections of t

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Reinforcement Learning for LLM-based Multi-Agent Systems through Orchestration Traces

DGX agent

arXiv:2605.02801v1 Announce Type: new Abstract: As large language model (LLM) agents evolve from isolated tool users into coordinated teams, reinforcement learning (RL) must optimize not only individu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Rethinking Multi-Label Node Classification: Do Tuned Classic GNNs Suffice?

DGX agent

arXiv:2605.01403v1 Announce Type: new Abstract: Multi-label node classification (MLNC) has recently been addressed by increasingly complex label-aware designs that explicitly model node-label interact

model-releasesarxiv-cs-lg
5 May 2026
Research

Rhamba: Region-Aware Hybrid Attention-Mamba Framework for Self-Supervised Learning in Resting-State fMRI

DGX agent

arXiv:2605.01240v1 Announce Type: new Abstract: Self-supervised pretraining is promising for large-scale neuroimaging, yet the impact of region-aware masking and hybrid sequence modeling remains under

researcharxiv-cs-lg
5 May 2026
Agents

S^3-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data

DGX agent

arXiv:2605.01248v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has enabled newer capabilities in models, such as agentic tool-use for search. However, these models struggle

agentsarxiv-cs-lg
5 May 2026
Model Releases

SAMamba3D: adapting Segment Anything for generalizable 3D segmentation of multiphase pore-scale images

DGX agent

arXiv:2605.00916v1 Announce Type: new Abstract: Reliable segmentation of multiphase pore-scale X-ray images of rocks is necessary to quantify fluid saturation, connectivity, and interfacial geometry.

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

SemEval-2026 Task 7: Everyday Knowledge Across Diverse Languages and Cultures

DGX agent

arXiv:2605.02601v1 Announce Type: new Abstract: We present our shared task on evaluating the adaptability of LLMs and NLP systems across multiple languages and cultures. The task data consist of an ex

model-releasesarxiv-cs-cl
5 May 2026
Safety

Skills as Verifiable Artifacts: A Trust Schema and a Biconditional Correctness Criterion for Human-in-the-Loop Agent Runtimes

DGX agent

arXiv:2605.00424v1 Announce Type: cross Abstract: Agent skills -- structured packages of instructions, scripts, and references that augment a large language model (LLM) without modifying the model its

safetyarxiv-cs-ai
5 May 2026
Model Releases

STEP: Warm-Started Visuomotor Policies with Spatiotemporal Consistency Prediction

DGX agent

arXiv:2602.08245v2 Announce Type: replace Abstract: Diffusion policies have recently emerged as a powerful paradigm for visuomotor control in robotic manipulation due to their ability to model the dis

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

SwiftChannel: Algorithm-Hardware Co-Design for Deep Learning-Based 5G Channel Estimation

DGX agent

arXiv:2605.01931v1 Announce Type: cross Abstract: Channel estimation is crucial in 5G communication networks for optimizing transmission parameters and ensuring reliable, high-speed communication. How

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Task-Driven Subspace Decomposition for Knowledge Sharing and Isolation in LoRA-based Continual Learning

DGX agent

arXiv:2603.00191v2 Announce Type: replace-cross Abstract: Continual Learning (CL) requires models to sequentially adapt to new tasks without forgetting old knowledge. Recently, Low-Rank Adaptation (Lo

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Teaching LLMs Brazilian Healthcare: Injecting Knowledge from Official Clinical Guidelines

DGX agent

arXiv:2605.01077v1 Announce Type: new Abstract: Brazil's Unified Health System (SUS) relies on official clinical guidelines that define diagnostic criteria, treatments, dosages, and monitoring procedu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

The Good, the Bad, and the Sampled: a No-Regret Approach to Safe Online Classification

DGX agent

arXiv:2510.01020v2 Announce Type: replace Abstract: We study sequential testing for a binary disease outcome when risk follows an unknown logistic model. At each round, the decision maker may either p

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Towards Lightest Low-Light Image Enhancement Architecture for Mobile Devices

DGX agent

arXiv:2507.04277v2 Announce Type: replace Abstract: Real-time low-light image enhancement on mobile and embedded devices requires models that balance visual quality and computational efficiency. Exist

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

VILAS: A VLA-Integrated Low-cost Architecture with Soft Grasping for Robotic Manipulation

DGX agent

arXiv:2605.02037v1 Announce Type: new Abstract: We present VILAS, a fully low-cost, modular robotic manipulation platform designed to support end-to-end vision-language-action (VLA) policy learning an

model-releasesarxiv-cs-ro
5 May 2026
Model Releases

Borrowed Geometry: Computational Reuse of Frozen Text-Pretrained Transformer Weights Across Modalities

DGX agent

arXiv:2605.00333v1 Announce Type: cross Abstract: Frozen Gemma 4 31B weights pretrained exclusively on text tokens, unmodified, transfer across modality boundaries through a thin trainable interface.

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection

DGX agent

arXiv:2605.00630v1 Announce Type: new Abstract: The proliferation of advanced AI video synthesis techniques poses an unprecedented challenge to digital video authenticity. Existing AI-generated video

model-releasesarxiv-cs-cv
4 May 2026
Research

Confidence Estimation in Automatic Short Answer Grading with LLMs

DGX agent

arXiv:2605.00200v1 Announce Type: new Abstract: Automatic Short Answer Grading (ASAG) with generative large language models (LLMs) has recently demonstrated strong performance without task-specific fi

researcharxiv-cs-cl
4 May 2026
Model Releases

Deepfakes: we need to re-think the concept of 'real' images

DGX agent

arXiv:2509.21864v2 Announce Type: replace Abstract: The wide availability and low usability barrier of modern image generation models has triggered the reasonable fear of criminal misconduct and negat

model-releasesarxiv-cs-cv
4 May 2026
Research

End-to-End Autoregressive Image Generation with 1D Semantic Tokenizer

DGX agent

arXiv:2605.00503v1 Announce Type: new Abstract: Autoregressive image modeling relies on visual tokenizers to compress images into compact latent representations. We design an end-to-end training pipel

researcharxiv-cs-cv
4 May 2026
Model Releases

FedACT: Concurrent Federated Intelligence across Heterogeneous Data Sources

DGX agent

arXiv:2605.00011v1 Announce Type: new Abstract: Federated Learning (FL) enables collaborative intelligence across decentralized data source devices in a privacy-preserving way. While substantial resea

model-releasesarxiv-cs-lg
4 May 2026
Model Releases

FinSafetyBench: Evaluating LLM Safety in Real-World Financial Scenarios

DGX agent

arXiv:2605.00706v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly applied in financial scenarios. However, they may produce harmful outputs, including facilitating illegal

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Firestore at Next '26: Unlock agentic development, search and MongoDB compatibility

DGX agent

In the era of AI agents, the distance between a big idea and a working application has never been shorter. As we lean more heavily on agents to help us build applications, a critical question remains:

model-releasesgoogle-cloud-ai
4 May 2026
← Previous
1…545546547548549…1371
Next →