AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,425 results
10 Apr 2026

Apple: Toward General Active Perception via Reinforcement Learning

Model ReleasesDGX agent

arXiv:2505.06182v5 Announce Type: replace-cross Abstract: Active perception is a fundamental skill that enables us humans to deal with uncertainty in our inherently partially observable environment. F

Attention Flows: Tracing LLM Conceptual Engagement via Story Summaries

SafetyDGX agent

arXiv:2604.06416v1 Announce Type: cross Abstract: Although LLM context lengths have grown, there is evidence that their ability to integrate information across long-form texts has not kept pace. We ev

AudioRole: An Audio Dataset for Character Role-Playing in Large Language Models

SafetyDGX agent

arXiv:2509.23435v2 Announce Type: replace-cross Abstract: The creation of high-quality multimodal datasets remains fundamental for advancing role-playing capabilities in large language models (LLMs).

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Beyond Facts: Benchmarking Distributional Reading Comprehension in Large Language Models

Model ReleasesDGX agent

arXiv:2604.06201v1 Announce Type: cross Abstract: While most reading comprehension benchmarks for LLMs focus on factual information that can be answered by localizing specific textual evidence, many r

Blockchain and AI: Securing Intelligent Networks for the Future

AgentsDGX agent

arXiv:2604.06323v2 Announce Type: cross Abstract: Blockchain and artificial intelligence (AI) are increasingly proposed together for securing intelligent networks, but the literature remains fragmente

Bridging Natural Language and Microgrid Dynamics: A Context-Aware Simulator and Dataset

ApplicationsDGX agent

arXiv:2604.05429v2 Announce Type: replace-cross Abstract: Addressing the critical need for intelligent, context-aware energy management in renewable systems, we introduce the OpenCEM Simulator and Dat

CAMotion: A High-Quality Benchmark for Camouflaged Moving Object Detection in the Wild

Model ReleasesDGX agent

arXiv:2604.08287v1 Announce Type: new Abstract: Discovering camouflaged objects is a challenging task in computer vision due to the high similarity between camouflaged objects and their surroundings.

Can Vision Language Models Judge Action Quality? An Empirical Evaluation

Model ReleasesDGX agent

arXiv:2604.08294v1 Announce Type: cross Abstract: Action Quality Assessment (AQA) has broad applications in physical therapy, sports coaching, and competitive judging. Although Vision Language Models

ChatGPT for marketing teams

TutorialsDGX agent

The OpenAI Academy's 'ChatGPT for Marketing Teams' tutorial provides marketing and brand teams with prompts and use-case guidance to streamline strategy, content creation, and performance analysis,...

Claude Mythos is too dangerous for public consumption...

Model ReleasesDGX agent

Anthropic announced **Claude Mythos Preview**, its most powerful AI model to date, which it is withholding from general public release due to its advanced and potentially dangerous cybersecurity ca...

Contextual Earnings-22: A Speech Recognition Benchmark with Custom Vocabulary in the Wild

Model ReleasesDGX agent

arXiv:2604.07354v1 Announce Type: new Abstract: The accuracy frontier of speech-to-text systems has plateaued on academic benchmarks.1 In contrast, industrial benchmarks and adoption in high-stakes do

CoreWeave inks multiyear cloud deal with Anthropic

Model ReleasesDGX agent

CoreWeave Inc. today announced that it has won a multiyear contract to supply Anthropic PBC with cloud infrastructure. The company’s shares closed 11% higher on the news. The data center capacity comm

Data Leakage in Automotive Perception: Practitioners' Insights

SafetyDGX agent

arXiv:2604.06899v1 Announce Type: cross Abstract: Data leakage is the inadvertent transfer of information between training and evaluation datasets that poses a subtle, yet critical, risk to the reliab

DinoRADE: Full Spectral Radar-Camera Fusion with Vision Foundation Model Features for Multi-class Object Detection in Adverse Weather

AgentsDGX agent

arXiv:2604.08074v1 Announce Type: new Abstract: Reliable and weather-robust perception systems are essential for safe autonomous driving and typically employ multi-modal sensor configurations to achie

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

SafetyDGX agent

arXiv:2604.07607v1 Announce Type: cross Abstract: Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human da

Face-D(^2)CL: Multi-Domain Synergistic Representation with Dual Continual Learning for Facial DeepFake Detection

Model ReleasesDGX agent

arXiv:2604.08159v1 Announce Type: new Abstract: The rapid advancement of facial forgery techniques poses severe threats to public trust and information security, making facial DeepFake detection a cri

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

Model ReleasesDGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On

Model ReleasesDGX agent

arXiv:2604.08526v1 Announce Type: new Abstract: Given a person and a garment image, virtual try-on (VTO) aims to synthesize a realistic image of the person wearing the garment, while preserving their

FORGE:Fine-grained Multimodal Evaluation for Manufacturing Scenarios

Model ReleasesDGX agent

arXiv:2604.07413v1 Announce Type: new Abstract: The manufacturing sector is increasingly adopting Multimodal Large Language Models (MLLMs) to transition from simple perception to autonomous execution,

From Ground Truth to Measurement: A Statistical Framework for Human Labeling

SafetyDGX agent

arXiv:2604.07591v1 Announce Type: cross Abstract: Supervised machine learning assumes that labeled data provide accurate measurements of the concepts models are meant to learn. Yet in practice, human

Front-End Ethics for Sensor-Fused Health Conversational Agents: An Ethical Design Space for Biometrics

SafetyDGX agent

arXiv:2604.06203v1 Announce Type: cross Abstract: The integration of continuous data from built-in sensors and Large Language Models (LLMs) has fueled a surge of 'Sensor-Fused LLM agents' for personal

GameWorld: Towards Standardized and Verifiable Evaluation of Multimodal Game Agents

Model ReleasesDGX agent

arXiv:2604.07429v1 Announce Type: new Abstract: Towards an embodied generalist for real-world interaction, Multimodal Large Language Model (MLLM) agents still suffer from challenging latency, sparse f

GaussiAnimate: Reconstruct and Rig Animatable Categories with Level of Dynamics

ApplicationsDGX agent

arXiv:2604.08547v1 Announce Type: new Abstract: Free-form bones, that conform closely to the surface, can effectively capture non-rigid deformations, but lack a kinematic structure necessary for intui

Gaze to Insight: A Scalable AI Approach for Detecting Gaze Behaviours in Face-to-Face Collaborative Learning

ApplicationsDGX agent

arXiv:2604.03317v2 Announce Type: replace Abstract: Previous studies have illustrated the potential of analysing gaze behaviours in collaborative learning to provide educationally meaningful informati

Getting started with ChatGPT

TutorialsDGX agent

OpenAI Academy's 'Getting Started with ChatGPT' tutorial introduces users to ChatGPT as a conversational AI application built on large language models, teaching core concepts such as how to write e...

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models

HardwareDGX agent

arXiv:2604.07812v1 Announce Type: new Abstract: In multimodal large language models (MLLMs), the surge of visual tokens significantly increases the inference time and computational overhead, making th

Horticultural Temporal Fruit Monitoring via 3D Instance Segmentation and Re-Identification using Colored Point Clouds

ApplicationsDGX agent

arXiv:2411.07799v3 Announce Type: replace Abstract: Accurate and consistent fruit monitoring over time is a key step toward automated agricultural production systems. However, this task is inherently

How SAP Concur automates expense reporting with agentic AI

Model ReleasesDGX agent

For decades, expense automation relied on a simple premise: If the machine can read the text, it can do the work. But anyone who has ever tried to scan a crumpled, smudged, or sun-bleached receipt fro

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Mo…

Model ReleasesDGX agent

I strongly suspect that Claude Mythos is a looped language model, as described in the paper 'Scaling Latent Reasoning via Looped Language Models' from ByteDance The authors of that paper called out gr

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

Model ReleasesDGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

Knowledge Graphs Generation from Cultural Heritage Texts: Combining LLMs and Ontological Engineering for Scholarly Debates

Model ReleasesDGX agent

arXiv:2511.10354v1 Announce Type: cross Abstract: Cultural Heritage texts contain rich knowledge that is difficult to query systematically due to the challenges of converting unstructured discourse in

LLM-Based Data Generation and Clinical Skills Evaluation for Low-Resource French OSCEs

ApplicationsDGX agent

arXiv:2604.08126v1 Announce Type: new Abstract: Objective Structured Clinical Examinations (OSCEs) are the standard method for assessing medical students' clinical and communication skills through str

LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sources

SafetyDGX agent

arXiv:2604.06571v1 Announce Type: cross Abstract: Missing-person and child-safety investigations rely on heterogeneous case documents, including structured forms, bulletin-style posters, and narrative

LLM Spirals of Delusion: A Benchmarking Audit Study of AI Chatbot Interfaces

Model ReleasesDGX agent

arXiv:2604.06188v1 Announce Type: cross Abstract: People increasingly hold sustained, open-ended conversations with large language models (LLMs). Public reports and early studies suggest that, in such

LUMINA: Foundation Models for Topology Transferable ACOPF

SafetyDGX agent

arXiv:2603.04300v2 Announce Type: replace Abstract: Foundation models in general promise to accelerate scientific computation by learning reusable representations across problem instances, yet constra

Lyria by @GoogleDeepMind is next level it did average Turkish pop song very well all my Turkish friends in the conference hall was summoned …

ToolsDGX agent

Google DeepMind's Lyria is an advanced AI music generation model capable of producing high-quality audio — including vocals, lyrics, and instrumentals — across a wide range of genres and styles fro...

Matrix Profile for Anomaly Detection on Multidimensional Time Series

Model ReleasesDGX agent

arXiv:2409.09298v2 Announce Type: replace-cross Abstract: The Matrix Profile (MP), a versatile tool for time series data mining, has been shown effective in time series anomaly detection (TSAD). This

MedRoute: RL-Based Dynamic Specialist Routing in Multi-Agent Medical Diagnosis

AgentsDGX agent

arXiv:2604.06180v1 Announce Type: cross Abstract: Medical diagnosis using Large Multimodal Models (LMMs) has gained increasing attention due to capability of these models in providing precise diagnose

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

SafetyDGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

Mining Electronic Health Records to Investigate Effectiveness of Ensemble Deep Clustering

ApplicationsDGX agent

arXiv:2604.07085v1 Announce Type: new Abstract: In electronic health records (EHRs), clustering patients and distinguishing disease subtypes are key tasks to elucidate pathophysiology and aid clinical

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

Model ReleasesDGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

MM-MoralBench: A MultiModal Moral Evaluation Benchmark for Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2412.20718v2 Announce Type: replace Abstract: The rapid integration of Large Vision-Language Models (LVLMs) into critical domains necessitates comprehensive moral evaluation to ensure their alig

MolmoWeb: Open Visual Web Agent and Open Data for the Open Web

AgentsDGX agent

arXiv:2604.08516v1 Announce Type: new Abstract: Web agents--autonomous systems that navigate and execute tasks on the web on behalf of users--have the potential to transform how people interact with t

Oldest octopus fossil found to not be an octopus

IndustryDGX agent

A 300-million-year-old fossil named *Pohlsepia mazonensis*, originally identified as the world's oldest octopus in 2000, has been reclassified in 2026 as a nautiloid — a relative of the modern naut...

ORACLE-SWE: Quantifying the Contribution of Oracle Information Signals on SWE Agents

AgentsDGX agent

arXiv:2604.07789v1 Announce Type: cross Abstract: Recent advances in language model (LM) agents have significantly improved automated software engineering (SWE). Prior work has proposed various agenti

ParkSense: Where Should a Delivery Driver Park? Leveraging Idle AV Compute and Vision-Language Models

AgentsDGX agent

arXiv:2604.07912v1 Announce Type: new Abstract: Finding parking consumes a disproportionate share of food delivery time, yet no system addresses precise parking-spot selection relative to merchant ent

Personalizing Text-to-Image Generation to Individual Taste

SafetyDGX agent

arXiv:2604.07427v1 Announce Type: new Abstract: Modern text-to-image (T2I) models generate high-fidelity visuals but remain indifferent to individual user preferences. While existing reward models opt

PhyAVBench: A Challenging Audio Physics-Sensitivity Benchmark for Physically Grounded Text-to-Audio-Video Generation

Model ReleasesDGX agent

arXiv:2512.23994v3 Announce Type: replace-cross Abstract: Text-to-audio-video (T2AV) generation is central to applications such as filmmaking and world modeling. However, current models often fail to

PIKA: Expert-Level Synthetic Datasets for Post-Training Alignment from Scratch

Model ReleasesDGX agent

arXiv:2510.06670v2 Announce Type: replace Abstract: High-quality instruction data is critical for LLM alignment, yet existing open-source datasets often lack efficiency, requiring hundreds of thousand

Pistachio: Towards Synthetic, Balanced, and Long-Form Video Anomaly Benchmarks

Model ReleasesDGX agent

arXiv:2511.19474v4 Announce Type: replace-cross Abstract: Automatically detecting abnormal events in videos is crucial for modern autonomous systems, yet existing Video Anomaly Detection (VAD) benchma

Prompt reinforcing for long-term planning of large language models

Model ReleasesDGX agent

arXiv:2510.05921v2 Announce Type: replace Abstract: Large language models (LLMs) have achieved remarkable success in a wide range of natural language processing tasks and can be adapted through prompt

Quantitative Estimation of Target Task Performance from Unsupervised Pretext Task in Semi/Self-Supervised Learning

Model ReleasesDGX agent

arXiv:2508.07299v2 Announce Type: replace-cross Abstract: The effectiveness of unlabeled data in Semi/Self-Supervised Learning (SSL) depends on appropriate assumptions for specific scenarios, thereby

SceneScribe-1M: A Large-Scale Video Dataset with Comprehensive Geometric and Semantic Annotations

Model ReleasesDGX agent

arXiv:2604.07990v1 Announce Type: new Abstract: The convergence of 3D geometric perception and video synthesis has created an unprecedented demand for large-scale video data that is rich in both seman

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

Model ReleasesDGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

SearchAD: Large-Scale Rare Image Retrieval Dataset for Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.08008v1 Announce Type: new Abstract: Retrieving rare and safety-critical driving scenarios from large-scale datasets is essential for building robust autonomous driving (AD) systems. As dat

See where you can catch us next: https://www.together.ai/events

ToolsDGX agent

Together AI maintains a public events page at [together.ai/events](https://www.together.ai/events) listing its past and upcoming appearances, spanning major industry conferences such as NVIDIA GTC,...

Should we be optimizing for limited compute instead of more parameters? Thoughts?

Local AiDGX agent

"The search did not return the specific Reddit thread. However, I can provide a summary based on what the topic is broadly about within the local-AI/Ollama community context:

SMFD-UNet: Semantic Face Mask Is The Only Thing You Need To Deblur Faces

ApplicationsDGX agent

arXiv:2604.07477v1 Announce Type: new Abstract: For applications including facial identification, forensic analysis, photographic improvement, and medical imaging diagnostics, facial image deblurring

Steering the Verifiability of Multimodal AI Hallucinations

TutorialsDGX agent

arXiv:2604.06714v1 Announce Type: new Abstract: AI applications driven by multimodal large language models (MLLMs) are prone to hallucinations and pose considerable risks to human users. Crucially, su

Tabular GANs for uneven distribution

Model ReleasesDGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

← Previous
1…421422423424
Next →