AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,424 results
Agents

Warmth and Competence in the Swarm: Designing Effective Human-Robot Teams

DGX agent

arXiv:2604.19270v1 Announce Type: new Abstract: As groups of robots increasingly collaborate with humans, understanding how humans perceive them is critical for designing effective human-robot teams.

agentsarxiv-cs-ro
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

We are excited to launch VideoGameBench on Antim Labs, created by @a1zhang, Thomas L. Griffiths (@cocosci_lab), @karthik_r_n, and @OfirPress…

DGX agent

VideoGameBench is a new benchmark launched on Antim Labs, created by a1zhang, Thomas L. Griffiths, Karthik R. N, and Ofir Press. The benchmark likely evaluates AI model performance on video game-relat

agentsyohei-nakajima--x
22 Apr 2026
Model Releases

What’s new in the Agentic Data Cloud: Powering the System of Action

DGX agent

Companies are shifting from gen AI that simply answers questions to autonomous agents that perceive, reason, and act on their behalf. Attempting to scale these agents on legacy stacks exposes structur

model-releasesgoogle-cloud-ai
22 Apr 2026
Model Releases

What’s next in Google AI infrastructure: Scaling for the agentic era

DGX agent

AI is evolving from answering questions to reasoning and taking action. Companies who want to lead in today’s agentic era require computing infrastructure designed and optimized for these new requirem

model-releasesgoogle-cloud-ai
22 Apr 2026
Agents

Advancing MAPF Toward the Real World: A Scalable Multi-Agent Realistic Testbed (SMART)

DGX agent

arXiv:2503.04798v3 Announce Type: replace Abstract: We present Scalable Multi-Agent Realistic Testbed (SMART), a realistic and efficient software tool for evaluating Multi-Agent Path Finding (MAPF) al

agentsarxiv-cs-ro
21 Apr 2026
Industry

AI Insider: 'Adding a Human Makes Your Team Worse' Emad Mostaque | @EMostaque TIMESTAMPS : 00:00 The models too dangerous to release 10:06 W…

DGX agent

AI Insider: 'Adding a Human Makes Your Team Worse' Emad Mostaque | @EMostaque TIMESTAMPS : 00:00 The models too dangerous to release 10:06 Why physics needs axioms — and AI doesn't 11:43 The MIND fram

industryemad-mostaque--x
21 Apr 2026
Model Releases

AIM 2025 Rip Current Segmentation (RipSeg) Challenge Report

DGX agent

arXiv:2508.13401v3 Announce Type: replace Abstract: This report presents an overview of the AIM 2025 RipSeg Challenge, a competition designed to advance techniques for automatic rip current segmentati

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning

DGX agent

arXiv:2604.16622v1 Announce Type: new Abstract: Backchannels (e.g., `yeah', `mhm', and `right') are short, non-interruptive feedback signals whose lexical form and prosody jointly convey pragmatic mea

safetyarxiv-cs-cl
21 Apr 2026
Applications

AlphaContext: An Evolutionary Tree-based Psychometric Context Generator for Creativity Assessment

DGX agent

arXiv:2604.18398v1 Announce Type: new Abstract: Creativity has become a core competence in the era of LLMs and human-AI collaboration, underpinning innovation in real-world problem solving. Crucially,

applicationsarxiv-cs-cl
21 Apr 2026
Safety

Annotation-Assisted Learning of Treatment Policies From Multimodal Electronic Health Records

DGX agent

arXiv:2507.20993v3 Announce Type: replace Abstract: We study how to learn treatment policies from multimodal electronic health records (EHRs) that consist of tabular data and clinical text. These poli

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Anthropic gets $5B investment from Amazon, will use it to buy Amazon chips

DGX agent

Amazon announced a 5 billion investment in Anthropic, with up to 20 billion more tied to commercial milestones. Anthropic committed to spending over $100 billion on AWS technologies over the next deca

model-releasesars-technica
21 Apr 2026
Model Releases

Automatic Dataset Construction (ADC): Sample Collection, Data Curation, and Beyond

DGX agent

arXiv:2408.11338v2 Announce Type: replace-cross Abstract: Large-scale data collection is essential for developing personalized training data, mitigating the shortage of training data, and fine-tuning

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Automatic Slide Updating with User-Defined Dynamic Templates and Natural Language Instructions

DGX agent

arXiv:2604.17894v1 Announce Type: new Abstract: Presentation slides are a primary medium for data-driven reporting, yet keeping complex, analytics-style decks up to date remains labor-intensive. Exist

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

BenchMarker: An Education-Inspired Toolkit for Highlighting Flaws in Multiple-Choice Benchmarks

DGX agent

arXiv:2602.06221v2 Announce Type: replace Abstract: Multiple-choice question answering (MCQA) is standard in NLP, but benchmarks lack rigorous quality control. We present BenchMarker, an education-ins

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

BIASEDTALES-ML: A Multilingual Dataset for Analyzing Narrative Attribute Distributions in LLM-Generated Stories

DGX agent

arXiv:2604.17008v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used to generate narrative content, including children's stories, which play an important role in social a

safetyarxiv-cs-cl
21 Apr 2026
Applications

CanonSLR: Canonical-View Guided Multi-View Continuous Sign Language Recognition

DGX agent

arXiv:2604.18184v1 Announce Type: new Abstract: Continuous Sign Language Recognition (CSLR) has achieved remarkable progress in recent years; however, most existing methods are developed under single-

applicationsarxiv-cs-cv
21 Apr 2026
Safety

CAPC-CG: A Large-Scale, Expert-Directed LLM-Annotated Corpus of Adaptive Policy Communication in China

DGX agent

arXiv:2510.08986v2 Announce Type: replace Abstract: We introduce CAPC-CG, the Chinese Adaptive Policy Communication (Central Government) Corpus, the first open dataset of Chinese policy directives ann

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Capture Timing-Attention of Events in Clinical Time Series

DGX agent

arXiv:2602.10385v2 Announce Type: replace Abstract: Automatically discovering personalized sequential events from large-scale time-series data is crucial for enabling precision medicine in clinical re

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval

DGX agent

arXiv:2601.17230v2 Announce Type: replace Abstract: Automated Fact-Checking has largely focused on verifying general knowledge against static corpora, overlooking high-stakes domains like law where tr

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CATP: Confidence-Aware Token Pruning for Camouflaged Object Detection

DGX agent

arXiv:2604.16854v1 Announce Type: new Abstract: Camouflaged Object Detection (COD) aims to segment targets that share extreme textural and structural similarities with their complex environments. Leve

model-releasesarxiv-cs-cv
21 Apr 2026
Applications

CBR-to-SQL: Rethinking Retrieval-based Text-to-SQL using Case-based Reasoning in the Healthcare Domain

DGX agent

arXiv:2603.05569v2 Announce Type: replace-cross Abstract: Extracting insights from Electronic Health Record (EHR) databases often requires SQL expertise, creating a barrier for clinical decision-makin

applicationsarxiv-cs-cl
21 Apr 2026
Model Releases

CFMS: Towards Explainable and Fine-Grained Chinese Multimodal Sarcasm Detection Benchmark

DGX agent

arXiv:2604.16372v1 Announce Type: new Abstract: Multimodal sarcasm detection has recently garnered significant attention. However, existing benchmarks suffer from coarse-grained annotations and limite

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CORP: A Multi-Modal Dataset for Campus-Oriented Roadside Perception Tasks

DGX agent

arXiv:2404.03191v3 Announce Type: replace Abstract: Numerous roadside perception datasets have been introduced to propel advancements in autonomous driving and intelligent transportation systems resea

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

CRISP: Compressing Redundancy in Chain-of-Thought via Intrinsic Saliency Pruning

DGX agent

arXiv:2604.17297v1 Announce Type: new Abstract: Long Chain-of-Thought (CoT) reasoning is pivotal for the success of recent reasoning models but suffers from high computational overhead and latency. Wh

safetyarxiv-cs-cl
21 Apr 2026
Safety

Decoding AI Tutor Effects for Educational Measurement: Temporal, Multi-Outcome, and Behavior-Cognitive Analysis

DGX agent

arXiv:2604.16366v1 Announce Type: cross Abstract: Artificial intelligence (AI) tutors have become increasingly popular in learning environments. In this study, we propose an AI agent prototype framewo

safetyarxiv-cs-lg
21 Apr 2026
Safety

Differential Privacy in Two-Layer Networks: How DP-SGD Harms Fairness and Robustness

DGX agent

arXiv:2603.04881v2 Announce Type: replace Abstract: Differentially private learning is essential for training models on sensitive data, but empirical studies consistently show that it can degrade perf

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

DSH-Bench: A Difficulty- and Scenario-Aware Benchmark with Hierarchical Subject Taxonomy for Subject-Driven Text-to-Image Generation

DGX agent

arXiv:2603.08090v2 Announce Type: replace Abstract: Significant progress has been achieved in subject-driven text-to-image (T2I) generation, which aims to synthesize new images depicting target subjec

model-releasesarxiv-cs-cv
21 Apr 2026
Agents

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation

DGX agent

arXiv:2604.17473v1 Announce Type: new Abstract: Vision-Language Navigation(VLN) requires an agent to navigate through 3D environments by following natural language instructions. While recent Video Lar

agentsarxiv-cs-cv
21 Apr 2026
Applications

EgoWalk: A Multimodal Dataset for Robot Navigation in the Wild

DGX agent

arXiv:2505.21282v2 Announce Type: replace Abstract: Data-driven navigation algorithms are critically dependent on large-scale, high-quality real-world data collection for successful training and robus

applicationsarxiv-cs-ro
21 Apr 2026
Model Releases

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning

DGX agent

arXiv:2508.07809v5 Announce Type: replace Abstract: Reinforcement learning with verifiable reward (RLVR) has become a promising paradigm for post-training large language models (LLMs) to improve their

model-releasesarxiv-cs-lg
21 Apr 2026
Tutorials

Expert-Annotated Embryo Image Dataset with Natural Language Descriptions for Evidence-Based Patient Communication in IVF

DGX agent

arXiv:2604.16528v1 Announce Type: new Abstract: Embryo selection is one of multiple crucial steps in in-vitro fertilization, commonly based on morphological assessment by clinical embryologists. Altho

tutorialsarxiv-cs-cv
21 Apr 2026
Safety

Fairness Constraints in High-Dimensional Generalized Linear Models

DGX agent

arXiv:2604.16610v1 Announce Type: cross Abstract: Machine learning models often inherit biases from historical data, raising critical concerns about fairness and accountability. Conventional fairness

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready t…

DGX agent

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready to fit the formats you need, from wide banners and presentati

model-releasesopenai--x
21 Apr 2026
Agents

From Clinical Intent to Clinical Model: An Autonomous Coding-Agent Framework for Clinician-driven AI Development

DGX agent

arXiv:2604.17110v1 Announce Type: new Abstract: Clinical AI development has traditionally followed a collaborative paradigm that depends on close interaction between clinicians and specialized AI team

agentsarxiv-cs-cv
21 Apr 2026
Applications

From Static Inference to Dynamic Interaction: A Survey of Streaming Large Language Models

DGX agent

arXiv:2603.04592v3 Announce Type: replace Abstract: Standard Large Language Models (LLMs) are predominantly designed for static inference with pre-defined inputs, which limits their applicability in d

applicationsarxiv-cs-cl
21 Apr 2026
Agents

HeLa-Mem: Hebbian Learning and Associative Memory for LLM Agents

DGX agent

arXiv:2604.16839v1 Announce Type: new Abstract: Long-term memory is a critical challenge for Large Language Model agents, as fixed context windows cannot preserve coherence across extended interaction

agentsarxiv-cs-cl
21 Apr 2026
Model Releases

HORIZON: A Benchmark for In-the-wild User Behaviour Modeling

DGX agent

arXiv:2604.17259v1 Announce Type: cross Abstract: User behavior in the real world is diverse, cross-domain, and spans long time horizons. Existing user modeling benchmarks however remain narrow, focus

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study

DGX agent

arXiv:2505.15404v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on reasoning-intensive tasks such as mathematics and programming. However, their enha

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HQA-VLAttack: Towards High Quality Adversarial Attack on Vision-Language Pre-Trained Models

DGX agent

arXiv:2604.16499v1 Announce Type: new Abstract: Black-box adversarial attack on vision-language pre-trained models is a practical and challenging task, as text and image perturbations need to be consi

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's W…

DGX agent

I came up with a somewhat foolish new benchmark for testing image generation models, to exercise the new ChatGPT Images 2.0: 'Do a where's Waldo style image but it's where is the raccoon holding a ham

model-releasessimon-willison--x
21 Apr 2026
Agents

Instruction-as-State: Environment-Guided and State-Conditioned Semantic Understanding for Embodied Navigation

DGX agent

arXiv:2604.18223v1 Announce Type: new Abstract: Vision-and-Language Navigation requires agents to follow natural-language instructions in visually changing environments. A central challenge is the dyn

agentsarxiv-cs-cv
21 Apr 2026
Applications

Integrating Feature Selection and Machine Learning for Nitrogen Assessment in Grapevine Leaves using In-Field Hyperspectral Imaging

DGX agent

arXiv:2507.17869v3 Announce Type: replace-cross Abstract: Nitrogen (N) is one of the most critical nutrients in winegrape production, influencing vine vigor, fruit composition, and wine quality. Becau

applicationsarxiv-cs-cv
21 Apr 2026
Model Releases

INTENT: Invariance and Discrimination-aware Noise Mitigation for Robust Composed Image Retrieval

DGX agent

arXiv:2604.18051v1 Announce Type: new Abstract: Composed Image Retrieval (CIR) is a challenging image retrieval paradigm that enables to retrieve target images based on multimodal queries consisting o

model-releasesarxiv-cs-cv
21 Apr 2026
Industry

it's entirely fine and great for proprietary models to exist the issue is they try to kill open source behind scenes while pretending they s…

DGX agent

it's entirely fine and great for proprietary models to exist the issue is they try to kill open source behind scenes while pretending they support it so far they've been pretty incompetent but they'll

industryclem-delangue--x
21 Apr 2026
Safety

Jailbreaking Large Language Models with Morality Attacks

DGX agent

arXiv:2604.17053v1 Announce Type: new Abstract: Pluralism alignment with AI has the sophisticated and necessary goal of creating AI that can coexist with and serve morally multifaceted humanity. Resea

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription

DGX agent

arXiv:2502.20295v2 Announce Type: replace-cross Abstract: Handwriting text recognition (HTR) remains a challenging task. Existing approaches require fine-tuning on labeled data, which is impractical t

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Learned Nonlocal Feature Matching and Filtering for RAW Image Denoising

DGX agent

arXiv:2604.17453v1 Announce Type: cross Abstract: Being one of the oldest and most basic problems in image processing, image denoising has seen a resurgence spurred by rapid advances in deep learning.

model-releasesarxiv-cs-cv
21 Apr 2026
Agents

Learning to Trade Like an Expert: Cognitive Fine-Tuning for Stable Financial Reasoning in Language Models

DGX agent

arXiv:2604.16862v1 Announce Type: new Abstract: Recent deployments of large language models (LLMs) as autonomous trading agents raise questions about whether financial decision-making competence gener

agentsarxiv-cs-lg
21 Apr 2026
← Previous
1…517518519520521…530
Next →