AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,429 results
21 Apr 2026

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready t…

Model ReleasesDGX agent

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready to fit the formats you need, from wide banners and presentati

From Clinical Intent to Clinical Model: An Autonomous Coding-Agent Framework for Clinician-driven AI Development

AgentsDGX agent

arXiv:2604.17110v1 Announce Type: new Abstract: Clinical AI development has traditionally followed a collaborative paradigm that depends on close interaction between clinicians and specialized AI team

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models

Applications
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2603.04592v3 Announce Type: replace Abstract: Standard Large Language Models (LLMs) are predominantly designed for static inference with pre-defined inputs, which limits their applicability in d

HeLa-Mem: Hebbian Learning and Associative Memory for LLM Agents

AgentsDGX agent

arXiv:2604.16839v1 Announce Type: new Abstract: Long-term memory is a critical challenge for Large Language Model agents, as fixed context windows cannot preserve coherence across extended interaction

HORIZON: A Benchmark for In-the-wild User Behaviour Modeling

Model ReleasesDGX agent

arXiv:2604.17259v1 Announce Type: cross Abstract: User behavior in the real world is diverse, cross-domain, and spans long time horizons. Existing user modeling benchmarks however remain narrow, focus

How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study

Model ReleasesDGX agent

arXiv:2505.15404v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on reasoning-intensive tasks such as mathematics and programming. However, their enha

HQA-VLAttack: Towards High Quality Adversarial Attack on Vision-Language Pre-Trained Models

Model ReleasesDGX agent

arXiv:2604.16499v1 Announce Type: new Abstract: Black-box adversarial attack on vision-language pre-trained models is a practical and challenging task, as text and image perturbations need to be consi

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's W…

Model ReleasesDGX agent

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's Waldo style image but it's where is the raccoon holding a ham

Instruction-as-State: Environment-Guided and State-Conditioned Semantic Understanding for Embodied Navigation

AgentsDGX agent

arXiv:2604.18223v1 Announce Type: new Abstract: Vision-and-Language Navigation requires agents to follow natural-language instructions in visually changing environments. A central challenge is the dyn

Integrating Feature Selection and Machine Learning for Nitrogen Assessment in Grapevine Leaves using In-Field Hyperspectral Imaging

ApplicationsDGX agent

arXiv:2507.17869v3 Announce Type: replace-cross Abstract: Nitrogen (N) is one of the most critical nutrients in winegrape production, influencing vine vigor, fruit composition, and wine quality. Becau

INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval

Model ReleasesDGX agent

arXiv:2604.18051v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) is a challenging image retrieval paradigm that enables to retrieve target images based on multimodal queries consisting o

it's entirely fine and great for proprietary models to exist the issue is they try to kill open source behind scenes while pretending they s…

IndustryDGX agent

it's entirely fine and great for proprietary models to exist the issue is they try to kill open source behind scenes while pretending they support it so far they've been pretty incompetent but they'll

Jailbreaking Large Language Models with Morality Attacks

SafetyDGX agent

arXiv:2604.17053v1 Announce Type: new Abstract: Pluralism alignment with AI has the sophisticated and necessary goal of creating AI that can coexist with and serve morally multifaceted humanity. Resea

Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription

Model ReleasesDGX agent

arXiv:2502.20295v2 Announce Type: replace-cross Abstract: Handwriting text recognition (HTR) remains a challenging task. Existing approaches require fine-tuning on labeled data, which is impractical t

Learned Nonlocal Feature Matching and Filtering for RAW Image Denoising

Model ReleasesDGX agent

arXiv:2604.17453v1 Announce Type: cross Abstract: Being one of the oldest and most basic problems in image processing, image denoising has seen a resurgence spurred by rapid advances in deep learning.

Learning to Trade Like an Expert: Cognitive Fine-Tuning for Stable Financial Reasoning in Language Models

AgentsDGX agent

arXiv:2604.16862v1 Announce Type: new Abstract: Recent deployments of large language models (LLMs) as autonomous trading agents raise questions about whether financial decision-making competence gener

Less Noise, More Voice: Reinforcement Learning for Reasoning via Instruction Purification

SafetyDGX agent

arXiv:2601.21244v3 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has advanced LLM reasoning, but remains constrained by inefficient exploration under lim

Leveraging Large Language Models for Sarcastic Speech Annotation in Sarcasm Detection

Model ReleasesDGX agent

arXiv:2506.00955v2 Announce Type: replace Abstract: Sarcasm fundamentally alters meaning through tone and context, yet detecting it in speech remains a challenge due to data scarcity. In addition, exi

Loneliness in older adults can often lead to memory impairment

IndustryDGX agent

A major European study tracking over 10,000 people aged 65-94 for seven years found that lonely older adults performed worse on initial memory tests, though their memory declined at the same rate as t

LookasideVLN: Direction-Aware Aerial Vision-and-Language Navigation

AgentsDGX agent

arXiv:2604.17190v1 Announce Type: new Abstract: Aerial Vision-and-Language Navigation (Aerial VLN) enables unmanned aerial vehicles (UAVs) to follow natural language instructions and navigate complex

MeasHalu: Mitigation of Scientific Measurement Hallucinations for Large Language Models with Enhanced Reasoning

Model ReleasesDGX agent

arXiv:2604.16929v1 Announce Type: new Abstract: The accurate extraction of scientific measurements from literature is a critical yet challenging task in AI4Science, enabling large-scale analysis and i

Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training (Reuters)

SafetyDGX agent

Reuters: Meta is installing tracking software on US staffers' computers to capture mouse movements, clicks, and keystrokes in work-related apps for use in AI training — Meta (META.O) is installing new

Mind the Way You Select Negative Texts: Pursuing the Distance Consistency in OOD Detection with VLMs

Model ReleasesDGX agent

arXiv:2603.02618v3 Announce Type: replace Abstract: Out-of-distribution (OOD) detection seeks to identify samples from unknown classes, a critical capability for deploying machine learning models in o

MM-Hand: A 21-DOF Multi-modal Modular Dexterous Robotic Hand with Remote Actuation

TutorialsDGX agent

arXiv:2604.17245v1 Announce Type: new Abstract: High-DOF dexterous hands require compact actuation, rich sensing, and reliable thermal behavior, but conventional designs often occupy valuable in-hand

// Multi-Agent Synthesis RAG // Nice paper on improving RAG systems with multiple agents. (bookmark it) The paper introduces MASS-RAG, a mul…

AgentsDGX agent

// Multi-Agent Synthesis RAG // Nice paper on improving RAG systems with multiple agents. (bookmark it) The paper introduces MASS-RAG, a multi-agent synthesis framework for retrieval-augmented generat

Multilevel neural networks with dual-stage feature fusion for human activity recognition

Model ReleasesDGX agent

arXiv:2604.16577v1 Announce Type: new Abstract: Human activity recognition (HAR) refers to the process of identifying human actions and activities using data collected from sensors. Neural networks, s

Multimodal Policy Internalization for Conversational Agents

SafetyDGX agent

arXiv:2510.09474v2 Announce Type: replace Abstract: Modern conversational agents like ChatGPT and Alexa+ rely on predefined policies specifying metadata, response styles, and tool-usage rules. As thes

NaviFormer: A Deep Reinforcement Learning Transformer-like Model to Holistically Solve the Navigation Problem

ApplicationsDGX agent

arXiv:2604.16967v1 Announce Type: new Abstract: Path planning is usually solved by addressing either the (high-level) route planning problem (waypoint sequencing to achieve the final goal) or the (low

Navigating Distribution Shifts in Medical Image Analysis: A Survey

SafetyDGX agent

arXiv:2411.05824v3 Announce Type: replace-cross Abstract: Medical Image Analysis (MedIA) has become indispensable in modern healthcare, enhancing clinical diagnostics and personalized treatment. Despi

Neighbor Embedding for High-Dimensional Sparse Poisson Data

ApplicationsDGX agent

arXiv:2604.16932v1 Announce Type: cross Abstract: Across many scientific fields, measurements often represent the number of times an event occurs. For example, a document can be represented by word oc

NTIRE 2026 Rip Current Detection and Segmentation (RipDetSeg) Challenge Report

Model ReleasesDGX agent

arXiv:2604.17070v1 Announce Type: new Abstract: This report presents the NTIRE 2026 Rip Current Detection and Segmentation (RipDetSeg) Challenge, which targets automatic rip current understanding in i

OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation

Model ReleasesDGX agent

arXiv:2506.05606v5 Announce Type: replace Abstract: Can large language models (LLMs) accurately simulate the next web action of a specific user? While LLMs have shown promising capabilities in generat

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

SafetyDGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

OptunaHub: A Platform for Black-Box Optimization

Model ReleasesDGX agent

arXiv:2510.02798v2 Announce Type: replace Abstract: Black-box optimization (BBO) underpins advances in domains such as AutoML and Materials Informatics, yet implementations of algorithms and benchmark

PBSBench: A Multi-Level Vision-Language Framework and Benchmark for Hematopathology Whole Slide Image Interpretation

Model ReleasesDGX agent

arXiv:2604.17570v1 Announce Type: new Abstract: Peripheral Blood Smear (PBS) is a critical microscopic examination in hematopathology that yields whole-slide imaging (WSI). Unlike solid tissue patholo

people compared GPT-5.4's solution of erdos #1196 to alphago's move 37 but i think a tighter analogy is to alphago's unusual (at the time) p…

Model ReleasesDGX agent

people compared GPT-5.4's solution of erdos #1196 to alphago's move 37 but i think a tighter analogy is to alphago's unusual (at the time) preferences for the 3-3 opening and early 3-3 invasion agains

PFDelta: A Benchmark Dataset for Power Flow under Load, Generation, and Topology Variations

Model ReleasesDGX agent

arXiv:2510.22048v3 Announce Type: replace Abstract: Power flow (PF) calculations are the backbone of real-time grid operations, across workflows such as contingency analysis (where repeated PF evaluat

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks

Model ReleasesDGX agent

arXiv:2602.06663v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have shown impressive capabilities in generating natural images and supporting multimodal reasoning. However, their

Plasticity Loss in Deep Reinforcement Learning: A Survey

SafetyDGX agent

arXiv:2411.04832v3 Announce Type: replace-cross Abstract: Plasticity refers to a network's ability to adapt to changing data distributions, which is crucial for the successful training of deep reinfor

PrefixMemory-Tuning: Modernizing Prefix-Tuning by Decoupling the Prefix from Attention

Model ReleasesDGX agent

arXiv:2506.13674v3 Announce Type: replace Abstract: Parameter-Efficient Fine-Tuning (PEFT) methods have become crucial for rapidly adapting large language models (LLMs) to downstream tasks. Prefix-Tun

Pulse Shape Discrimination Algorithms: Survey and Benchmark

Model ReleasesDGX agent

arXiv:2508.02750v2 Announce Type: replace Abstract: This review presents a comprehensive survey and benchmark of pulse shape discrimination (PSD) algorithms for radiation detection, classifying nearly

PyEPO: A PyTorch-based End-to-End Predict-then-Optimize Library for Linear and Integer Programming

TutorialsDGX agent

arXiv:2206.14234v3 Announce Type: replace-cross Abstract: In deterministic optimization, it is typically assumed that all problem parameters are fixed and known. In practice, however, some parameters

R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation

SafetyDGX agent

arXiv:2506.07826v2 Announce Type: replace Abstract: Validating autonomous driving (AD) systems requires diverse and safety-critical testing, making photorealistic virtual environments essential. Tradi

Really excited for this week! Next up, we've got something to show you at 12 pm PT today.

IndustryDGX agent

Sam Altman announced on X that OpenAI had something to reveal at 12 pm PT on a specific day, generating anticipation about an upcoming product announcement or demonstration. The post suggests a schedu

ReasonEmbed: Enhanced Text Embeddings for Reasoning-Intensive Document Retrieval

Model ReleasesDGX agent

arXiv:2510.08252v2 Announce Type: replace-cross Abstract: In this paper, we introduce ReasonEmbed, a novel text embedding model developed for reasoning-intensive document retrieval. Our work includes

Rethinking Meeting Effectiveness: A Benchmark and Framework for Temporal Fine-grained Automatic Meeting Effectiveness Evaluation

Model ReleasesDGX agent

arXiv:2604.17260v1 Announce Type: new Abstract: Evaluating meeting effectiveness is crucial for improving organizational productivity. Current approaches rely on post-hoc surveys that yield a single c

ReTrack: Evidence-Driven Dual-Stream Directional Anchor Calibration Network for Composed Video Retrieval

Model ReleasesDGX agent

arXiv:2604.17898v1 Announce Type: new Abstract: With the rapid growth of video data, Composed Video Retrieval (CVR) has emerged as a novel paradigm in video retrieval and is receiving increasing atten

R&F-Inventory: A Large-Scale Dataset for Monotonic Inventory Estimation in Reach and Frequency Advertising

Model ReleasesDGX agent

arXiv:2604.16821v1 Announce Type: new Abstract: Reach and Frequency (R&F) contract advertising is an important form of widely used brand advertising. Unlike performance advertising, R&F contracts emph

Same Claim, Different Judgment: Benchmarking Scenario-Induced Bias in Multilingual Financial Misinformation Detection

Model ReleasesDGX agent

arXiv:2601.05403v2 Announce Type: replace Abstract: Large language models (LLMs) have been widely applied across various domains of finance. Since their training data are largely derived from human-au

SciDraw-6K: A Multilingual Scientific Illustration Dataset Generated by Google Gemini

Model ReleasesDGX agent

arXiv:2604.17206v1 Announce Type: new Abstract: We present SciDraw-6K, a curated dataset of 6,291 scientific illustrations synthesized by Google Gemini image-generation models, each paired with prompt

StealthGraph: Exposing Domain-Specific Risks in LLMs through Knowledge-Graph-Guided Harmful Prompt Generation

SafetyDGX agent

arXiv:2601.04740v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly applied in specialized domains such as finance and healthcare, where they introduce unique safety risk

Structure-Aware Diversity Pursuit as an AI Safety Strategy against Homogenization

SafetyDGX agent

arXiv:2601.06116v2 Announce Type: replace-cross Abstract: Generative AI models reproduce the biases in the training data and can further amplify them through mode collapse. We refer to the resulting h

The Impact of Off-Policy Training Data on Probe Generalisation

SafetyDGX agent

arXiv:2511.17408v4 Announce Type: replace-cross Abstract: Probing has emerged as a promising method for monitoring large language models (LLMs), enabling cheap inference-time detection of concerning b

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey

AgentsDGX agent

arXiv:2509.02547v5 Announce Type: replace-cross Abstract: The emergence of agentic reinforcement learning (Agentic RL) marks a paradigm shift from conventional reinforcement learning applied to large

Thinking… Generating… Livestreaming… https://openai.com/live/

Model ReleasesDGX agent

OpenAI announced a live event showcasing new capabilities or features, likely including extended thinking, content generation, and livestreaming functionalities. The event demonstrated how these featu

Toward Reusability of AI Models Using Dynamic Updates of AI Documentation

SafetyDGX agent

arXiv:2604.17626v1 Announce Type: cross Abstract: This work addresses the challenge of disseminating reusable artificial intelligence (AI) models accompanied by AI documentation (a.k.a., AI model card

Towards a Foundation-Model Paradigm for Aerodynamic Prediction in Three-dimensional Design

TutorialsDGX agent

arXiv:2604.18062v1 Announce Type: new Abstract: Accurate machine-learning models for aerodynamic prediction are essential for accelerating shape optimization, yet remain challenging to develop for com

Towards Real-World Document Parsing via Realistic Scene Synthesis and Document-Aware Training

Model ReleasesDGX agent

arXiv:2603.23885v3 Announce Type: replace Abstract: Document parsing has recently advanced with multimodal large language models (MLLMs) that directly map document images to structured outputs. Tradit

TransXion: A High-Fidelity Graph Benchmark for Realistic Anti-Money Laundering

Model ReleasesDGX agent

arXiv:2604.17420v1 Announce Type: new Abstract: Money laundering poses severe risks to global financial systems, driving the widespread adoption of machine learning for transaction monitoring. However

UniDomain: Pretraining a Unified PDDL Domain from Real-World Demonstrations for Generalizable Robot Task Planning

ApplicationsDGX agent

arXiv:2507.21545v3 Announce Type: replace Abstract: Robotic task planning in real-world environments requires reasoning over implicit constraints from language and vision. While LLMs and VLMs offer st

← Previous
1…414415416417418…424
Next →