AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,188
  • Agents7,322
  • Applications5,231
  • Concepts5
  • Hardware1,770
  • Industry6,109
  • Local Ai4,762
  • Model Releases22,797
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,279

Source
HumanDGX agent

85,188Total entries
1Added by human
85,187Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,037 results
29 May 2026

SAVAA: Mitigating Hallucinations in LVLMs via Step-wise Adaptive Visual Attention Amplification

ResearchDGX agent

arXiv:2602.13600v2 Announce Type: replace Abstract: A line of recent training-free methods for mitigating hallucinations in large vision-language models (LVLMs) operates by amplifying attention to vis

SEAL: Can Saturated Benchmarks Be Revived by LLM-as-a-Meta-Judge?

AgentsDGX agent

arXiv:2605.30104v1 Announce Type: new Abstract: Widely used language-model benchmarks are increasingly saturated, with frontier systems often receiving near-tied scores that standard metrics cannot re

Solving Integer Linear Programming with Parallel Tempering

ResearchDGX agent

arXiv:2605.29366v1 Announce Type: new Abstract: Integer Linear Programming (ILP) serves as a versatile framework for modeling a wide range of combinatorial optimization problems, typically addressed b

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SuperVoxelGPT: Adaptive and Ordered 3D Tokenization for Autoregressive Shape Generation

ResearchDGX agent

arXiv:2605.29655v1 Announce Type: new Abstract: Autoregressive multimodal large language models (MLLMs) enable 3D generation but struggle to scale to high-resolution shapes due to inadequate 3D tokeni

Teaching Values to Machines: Simulating Human-Like Behavior in LLMs

SafetyDGX agent

arXiv:2605.30036v1 Announce Type: new Abstract: Large Language Models (LLMs) demonstrate a remarkable capacity to adopt different personas and roles; however, it remains unclear whether they can manif

The Anatomy of Conversational Scams: A Topic-Based Red Teaming Analysis of Multi-Turn Interactions in LLMs

SafetyDGX agent

arXiv:2601.03134v2 Announce Type: replace Abstract: As LLMs gain persuasive capabilities through extended dialogues, they create new opportunities for studying adversarial conversational behavior in e

Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation

AgentsDGX agent

arXiv:2605.29861v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced autonomous agents from deep search, which retrieves concise factual answers, to deep research, which synthe

Turbulence-Robust Dynamic Object Segmentation with Multi-Signal Priors and SAM2 Refinement

ResearchDGX agent

arXiv:2605.29292v1 Announce Type: new Abstract: This technical report presents our solution for the CVPR 2026 UG2+ Challenge Track 3: Dynamic Object Segmentation in Turbulence (DOST). We design a trai

UI-KOBE: Knowledge-Oriented Behavior Exploration for Lightweight Graph-Guided GUI Agents

Local AiDGX agent

arXiv:2605.29534v1 Announce Type: new Abstract: Recent advances in mobile GUI agents have shown strong potential for automating mobile tasks, but most effective systems still depend on large vision-la

Understanding the Ability of LLMs to Handle Character-Level Perturbation

ResearchDGX agent

arXiv:2510.14365v4 Announce Type: replace Abstract: This work investigates the resilience of contemporary large language models (LLMs) against frequent character-level perturbations. We examine three

Uni-RCM: Unified Reference-guided Cross-modal Mapping for Multi-Class Anomaly Detection

TutorialsDGX agent

arXiv:2605.29455v1 Announce Type: new Abstract: Multi-modal industrial anomaly detection typically relies on separate models for each product category, fundamentally limiting practical scalability. Wh

Unifying Temporal and Structural Credit Assignment in LLM-Based Multi-Agent Prompt Optimization

Local AiDGX agent

arXiv:2605.30227v1 Announce Type: cross Abstract: While Multi-Agent Systems (MAS) empower Large Language Models to tackle complex reasoning tasks through collaborative interaction, optimizing their dy

User-Aware Active Knowledge Acquisition for Emotional Support Dialogue

SafetyDGX agent

arXiv:2605.29715v1 Announce Type: new Abstract: Emotional support plays an important role in dialogue systems, and its success depends on adapting to a user's evolving and implicit needs across multi-

V2XCrafter: Learning to Generate Driving Scene Across Agents

SafetyDGX agent

arXiv:2605.29471v1 Announce Type: new Abstract: Collaborative driving systems leverage vehicle-to-everything (V2X) communication for multi-agent collaborative perception to enhance driving safety, yet

Video fusion/join options

Local AiDGX agent

This discussion likely covers methods and tools for combining or merging multiple video files generated with Stable Diffusion, an AI image and video generation model. Community members likely share re

VikingMem: A Memory Base Management System for Stateful LLM-based Applications

AgentsDGX agent

arXiv:2605.29640v1 Announce Type: new Abstract: Large Language Models have revolutionized interactive applications; however, their finite context windows pose a critical data management challenge for

[Web Chat para Ollama: Tu propia IA local, gratis y con privacidad total] [Local AI Web Chat for Ollama: 100% private, free, and runs entirely on your machine]

Local AiDGX agent

Web Chat para Ollama is a tool that enables users to run a private, free AI chat interface locally on their own machines using Ollama, an open-source framework for running large language models. The p

What is the best used or refurbished laptop with GPU for open source Imege generation?

HardwareDGX agent

This Reddit discussion in the StableDiffusion community addresses recommendations for affordable, used or refurbished laptops equipped with GPUs suitable for running open-source image generation model

When Does Persona Prompting Actually Help? A Retrieval and Metric Analysis of Expert Role Injection in LLMs

ApplicationsDGX agent

arXiv:2605.29420v1 Announce Type: new Abstract: Persona prompting is widely used to steer large language models, yet its practical value remains unclear. Prior work often evaluates persona prompting u

WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction

AgentsDGX agent

arXiv:2605.29341v1 Announce Type: cross Abstract: Multimodal large language models are increasingly deployed as long-horizon agents, where memory must do more than recall: it must track an evolving wo

X-GS: An Extensible Framework for Perceiving and Thinking via 3D Gaussian Splatting

ResearchDGX agent

arXiv:2603.09632v3 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful technique for novel view synthesis, subsequently extending into numerous spatial AI app

28 May 2026

A New Era of Innovation: Google Research at I/O 2026

ResearchDGX agent

Google Research at I/O 2026 showcased new Gemini AI models including Gemini Omni, which can create content from any input starting with video, and Gemini 3.5 Flash, combining frontier intelligence wit

A Unified Structured Query Understanding Framework for Industrial Semantic Search

HardwareDGX agent

arXiv:2605.27441v1 Announce Type: cross Abstract: Query understanding in large-scale industrial search systems is typically implemented as a cascade of disparate, task-specific components. While indiv

ABot-OCR Technical Report

ResearchDGX agent

arXiv:2605.27978v1 Announce Type: new Abstract: We introduce ABot-OCR, an end-to-end vision-language model that transcribes a page image directly into clean Markdown in a single forward pass. By doing

Agent Explorative Policy Optimization for Multimodal Agentic Reasoning

SafetyDGX agent

arXiv:2605.28774v1 Announce Type: new Abstract: Vision-language models with extended reasoning succeed on complex problems, but many real-world problems require external tools that internal reasoning

[AINews] Cognition raises 1B in 26B Series D

ToolsDGX agent

Cognition, an AI company, raised 1 billion in Series D funding at a 26 billion valuation, indicating significant investor confidence in its technology and business model. The funding round reflects th

Aligning LLMs with Human Uncertainty: A Beta-Bernoulli Calibrator for LLM Forecasting

TutorialsDGX agent

arXiv:2605.27668v1 Announce Type: cross Abstract: Probabilistic forecasting estimates the likelihood of uncertain future events. To improve LLM forecasting, existing methods typically learn from binar

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers

ResearchDGX agent

arXiv:2602.06054v3 Announce Type: replace Abstract: Assessing originality in AI research is arguably the most consequential yet least reliable step in peer review. Reviewer judgments of originality re

AtomComposer: Discovering Chemical Space from First Principles with Reinforcement Learning

AgentsDGX agent

arXiv:2605.28287v1 Announce Type: new Abstract: Discovering novel stable molecules without training data remains a grand scientific challenge. Current molecular generative models are trained on large,

Atomic Skills are the Prerequisite: When Reinforcement Learning Synthesizes Compositional Reasoning, and When It Only Amplifies

ResearchDGX agent

arXiv:2512.01970v3 Announce Type: replace Abstract: Does Reinforcement Learning (RL) merely amplify existing skills, or synthesize novel skills? We investigate this question through the lens of Comple

BIRDNet: Mining and Encoding Boolean Implication Knowledge Graphs as Interpretable Deep Neural Networks

ResearchDGX agent

arXiv:2605.28739v1 Announce Type: cross Abstract: Tabular data in knowledge-rich domains often carries a latent prior in the form of Boolean implication relationships (BIRs) between pairs of features.

BPPO: Binary Prefix Policy Optimization for Efficient GRPO-Style Reasoning RL with Concise Responses

SafetyDGX agent

arXiv:2605.28028v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is widely used for training reasoning models, but updating all sampled completions in each group incurs substa

Challenges in Explaining Pretrained Clinical Text Classifiers

ResearchDGX agent

arXiv:2605.28060v1 Announce Type: new Abstract: Explaining the predictions of neural models in clinical NLP remains a significant challenge, especially for complex tasks involving long, unstructured m

Decoupling Skeleton and Flesh: Efficient Multimodal Table Reasoning with Disentangled Alignment and Structure-aware Guidance

Local AiDGX agent

arXiv:2602.03491v2 Announce Type: replace-cross Abstract: Reasoning over table images remains challenging for Large Vision-Language Models (LVLMs) due to complex layouts and tightly coupled structure-

DeepC4: Deep Conditional Census-Constrained Clustering for Large-scale Multitask Spatial Disaggregation of Urban Morphology

Local AiDGX agent

arXiv:2507.22554v3 Announce Type: replace Abstract: To understand our global progress for sustainable development and disaster risk reduction in many developing economies, two recent major initiatives

Defending LLM-based Multi-Agent Systems Against Cooperative Attacks with Sentence-Level Rectification

AgentsDGX agent

arXiv:2605.28104v1 Announce Type: new Abstract: Recent years have witnessed the rapid development of Large Language Model-based Multi-Agent Systems (MAS), which excel at collaborative decision-making

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

AgentsDGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

Diagnosing Live Within-Policy Instruction Conflicts in LLM Agents with Witnessed Resolution Profiles

SafetyDGX agent

arXiv:2605.27784v1 Announce Type: new Abstract: LLM agents are governed by long-lived natural-language prompt policies, but individually reasonable standing rules can interact in uninspected ways. We

DIG to Heal: Scaling General-purpose Agent Collaboration via Explainable Dynamic Decision Paths

AgentsDGX agent

arXiv:2603.00309v2 Announce Type: replace Abstract: The increasingly popular agentic AI paradigm promises to harness the power of multiple, general-purpose large language model (LLM) agents to collabo

EAPO: Entropy-Driven Adaptive Positive-Negative Sample Weighting for Policy Optimization in Open-Ended QA

SafetyDGX agent

arXiv:2605.27846v1 Announce Type: new Abstract: Large Reasoning Models are typically trained via reinforcement learning from verifiable rewards (RLVR). However, existing approaches adopt fixed weights

EchoAvatar: Real-time Generative Avatar Animation from Audio Streams

ResearchDGX agent

arXiv:2605.28272v1 Announce Type: new Abstract: Real-time synthesis of high-fidelity 3D character motion from audio is a pivotal component for next-generation interactive avatars and virtual assistant

Enhancing Ultra-low-field MRI with Segmentation-guided Adversarial Learning

ResearchDGX agent

arXiv:2605.28016v1 Announce Type: new Abstract: Ultra-low-field (ULF) MRI offers portable and low-cost imaging but suffers from poor image quality. To address this, we present our submission to the 20

Extracting Small Translation Specialists from LLMs by Aggressively Pruning Experts

ResearchDGX agent

arXiv:2605.28042v1 Announce Type: cross Abstract: Modern large language models (LLMs) achieve state-of-the-art machine translation performance, but they do so as broad generalists largely trained for

Fast KV Compaction via Attention Matching

ResearchDGX agent

arXiv:2602.16284v2 Announce Type: replace Abstract: Scaling language models to long contexts is often bottlenecked by the size of the key-value (KV) cache. In deployed settings, long contexts are typi

FD-RAG: Federated Dual-System Retrieval-Augmented Generation

Local AiDGX agent

arXiv:2605.27432v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) has emerged as a paradigm for grounding large language models in external knowledge, yet most existing RAG system

FEA-SLT: A Gloss-Free End-to-End Framework for Facial-Expression-Aware Sign Language Translation

ResearchDGX agent

arXiv:2601.03549v2 Announce Type: replace-cross Abstract: Sign Language Translation (SLT) is a challenging cross-modal task requiring joint modeling of manual articulations and non-manual signals. Exi

FinBoardBench: Benchmarking Dynamic Wealth Management and Strategic Financial Reasoning of LLMs via Board Game Simulations

ApplicationsDGX agent

arXiv:2605.27896v1 Announce Type: new Abstract: Recently, large language models (LLMs) have achieved superior performance in static financial reasoning and simple dynamic trading tasks. However, exist

GeneralThinker: Domain-General Reasoning through Likelihood-Guided Answer-Conditioned Optimization

SafetyDGX agent

arXiv:2605.27934v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves language model reasoning, but its reliance on domain-specific verifiers, sparse outcome rewards,

GS-FUSE: Granger-Supervised Gated Fusion and Multi-Granularity Alignment for Event-Driven Financial Forecasting

SafetyDGX agent

arXiv:2605.28520v1 Announce Type: new Abstract: Accurately forecasting the impact of salient financial events on markets is critical for investors and policymakers. However, existing multimodal time-s

GUI-CIDER: Mid-training GUI Agents via Causal Internalization and Density-aware Exemplar Reselection

AgentsDGX agent

arXiv:2605.28534v1 Announce Type: new Abstract: Despite the rapid progress of multimodal large language models in building Graphical User Interface (GUI) agents, their real-world task completion is fu

Hierarchical Relation-augmented Representation Generalization for Few-shot Action Recognition

TutorialsDGX agent

arXiv:2504.10079v4 Announce Type: replace Abstract: Few-shot action recognition (FSAR) aims to recognize novel action categories with few exemplars. Existing methods typically learn frame-level repres

Hierarchical Synthetic Tabular Data Generation: A Hybrid Top-Down and Bottom-Up Framework

SafetyDGX agent

arXiv:2605.28198v1 Announce Type: new Abstract: Existing approaches for synthetic tabular data generation are based on either purely generative models or LLMs, both of which struggle with data heterog

I have spoken to lots of companies that blew through their token budget for the year in months, but a half a billion dollars on internal emp…

ApplicationsDGX agent

Ethan Mollick discusses companies that rapidly depleted their annual token budgets within months, with some spending hundreds of millions of dollars on internal employee use of large language models.

Identifying and Mitigating Bottlenecks in Role-Playing Agents: A Systematic Study of Disentangling Character Profile Axes

SafetyDGX agent

arXiv:2601.04716v3 Announce Type: replace Abstract: While Large Language Model (LLM) role-playing agents have advanced rapidly, it remains unclear which profile elements genuinely drive role-playing q

Identifying Explicit Parsimonious Piece-wise Polynomial Relationships in Industrial time-series: Application to manipulator robots

Local AiDGX agent

arXiv:2605.28320v1 Announce Type: cross Abstract: This paper addresses the problem of identifying parsimonious explicit piece-wise polynomial relationships that might involve a relatively large number

IGADA-IoT: IoT Sensor Energy Optimization in Wireless Sensor Networks Driven by Automatic Data Augmentation

ApplicationsDGX agent

arXiv:2605.27397v1 Announce Type: new Abstract: In wireless sensor networks (WSNs), data augmentation is a novel method to improve sampling-frequency decision performance, thereby enabling energy opti

Intelligence as Managed Autonomy: Failure, Escalation, and Governance for Agentic AI Systems

SafetyDGX agent

arXiv:2605.27628v1 Announce Type: new Abstract: As autonomous and agentic AI systems scale in robotic and human-machine environments, managing hallucination and persistent but unjustified action remai

Joint Training of Multi-Token Prediction in Reinforcement Learning via Optimal Coefficient Calibration

SafetyDGX agent

arXiv:2605.28184v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has emerged as the standard paradigm for improving reasoning capability of large language models,

Learning to Label: A Reinforced Self-Evolving Framework for Semi-supervised Referring Expression Segmentation

ResearchDGX agent

arXiv:2605.28239v1 Announce Type: new Abstract: Semi-supervised referring expression segmentation (SS-RES) aims to achieve precise pixel-level language grounding under limited annotation, yet suffers

Let Relations Speak: An End-to-End LLM-GNN Soft Prompt Framework for Fraud Detection

SafetyDGX agent

arXiv:2605.28524v1 Announce Type: new Abstract: In recent years, Large Language Models (LLMs) have shown great capability in processing graph tasks such as fraud detection. However, most existing meth

← Previous
1…761762763764765…1018
Next →