AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
26,542 results
Model Releases

LLMSniffer: Detecting LLM-Generated Code via GraphCodeBERT and Supervised Contrastive Learning

DGX agent

arXiv:2604.16058v1 Announce Type: cross Abstract: The rapid proliferation of Large Language Models (LLMs) in software development has made distinguishing AI-generated code from human-written code a cr

model-releasesarxiv-cs-cl
20 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MM-Telco: Benchmarks and Multimodal Large Language Models for Telecom Applications

DGX agent

arXiv:2511.13131v2 Announce Type: replace Abstract: Large Language Models (LLMs) have emerged as powerful tools for automating complex reasoning and decision-making tasks. In telecommunications, they

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

Nice paper combining the strength of Skills and RAG. Most RAG systems retrieve on every query, whether the model needs help or not. This is …

DGX agent

Nice paper combining the strength of Skills and RAG. Most RAG systems retrieve on every query, whether the model needs help or not. This is wasteful when the model already knows the answer, and often

agentsdair-ai--x
20 Apr 2026
Model Releases

PixDLM: A Dual-Path Multimodal Language Model for UAV Reasoning Segmentation

DGX agent

arXiv:2604.15670v1 Announce Type: new Abstract: Reasoning segmentation has recently expanded from ground-level scenes to remote-sensing imagery, yet UAV data poses distinct challenges, including obliq

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

Puppets or partners? Governing cyborg propaganda in the digital public square

DGX agent

arXiv:2602.13088v2 Announce Type: replace-cross Abstract: The distinction between genuine grassroots activism and automated influence operations is collapsing. While contemporary policy debates priori

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

RedBench: A Universal Dataset for Comprehensive Red Teaming of Large Language Models

DGX agent

arXiv:2601.03699v2 Announce Type: replace Abstract: As large language models (LLMs) become integral to safety-critical applications, ensuring their robustness against adversarial prompts is paramount.

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

RoleConflictBench: A Benchmark of Role Conflict Scenarios for Evaluating LLMs' Contextual Sensitivity

DGX agent

arXiv:2509.25897v2 Announce Type: replace-cross Abstract: People often encounter role conflicts -- social dilemmas where the expectations of multiple roles clash and cannot be simultaneously fulfilled

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Seed1.8 Model Card: Towards Generalized Real-World Agency

DGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Sketch and Text Synergy: Fusing Structural Contours and Descriptive Attributes for Fine-Grained Image Retrieval

DGX agent

arXiv:2604.15735v1 Announce Type: cross Abstract: Fine-grained image retrieval via hand-drawn sketches or textual descriptions remains a critical challenge due to inherent modality gaps. While hand-dr

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

TabularMath: Understanding Math Reasoning over Tables with Large Language Models

DGX agent

arXiv:2505.19563v4 Announce Type: replace Abstract: Mathematical reasoning has long been a key benchmark for evaluating large language models. Although substantial progress has been made on math word

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Technically Love: The Evolution of Human-AI Romance Discourse on Reddit

DGX agent

arXiv:2604.15333v1 Announce Type: cross Abstract: Human-AI romantic relationships are increasingly common, yet little is understood about how public discourse around them emerges and shifts over time.

applicationsarxiv-cs-ai
20 Apr 2026
Safety

The AI industry insists they can manage the risks of superintelligence, but there are in fact zero widely agreed on or accepted solutions to…

DGX agent

The AI industry insists they can manage the risks of superintelligence, but there are in fact zero widely agreed on or accepted solutions to the problem of how one could even control something vastly

safetyconnor-leahy--x
20 Apr 2026
Model Releases

the Codex x @skybysoftware acquisition may have been one of the best @openai deals made in the last year. I've been waiting for 'real' compu…

DGX agent

the Codex x @skybysoftware acquisition may have been one of the best @openai deals made in the last year. I've been waiting for 'real' computer use since @romainhuet demoed the ChatGPT App with 4o Vis

model-releasesswyx--x
20 Apr 2026
Applications

The Crutch or the Ceiling? How Different Generations of LLMs Shape EFL Student Writings

DGX agent

arXiv:2604.15460v1 Announce Type: cross Abstract: The rapid evolution of Large Language Models (LLMs) has made them powerful tools for enhancing student writing. This study explores the extent and lim

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world…

DGX agent

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world trade, policy and controls work. And in some deeper sense h

model-releasessoumith-chintala--x
20 Apr 2026
Model Releases

The Relic Condition: When Published Scholarship Becomes Material for Its Own Replacement

DGX agent

arXiv:2604.16116v1 Announce Type: cross Abstract: We extracted the scholarly reasoning systems of two internationally prominent humanities and social science scholars from their published corpora alon

model-releasesarxiv-cs-ai
20 Apr 2026
Hardware

The threat of analytic flexibility in using large language models to simulate human data

DGX agent

arXiv:2509.13397v3 Announce Type: replace-cross Abstract: Social scientists are now using large language models to create 'silicon samples': synthetic datasets intended to stand in for human responden

hardwarearxiv-cs-ai
20 Apr 2026
Safety

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

DGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

safetyarxiv-cs-ai
20 Apr 2026
Safety

Towards Robust Endogenous Reasoning: Unifying Drift Adaptation in Non-Stationary Tuning

DGX agent

arXiv:2604.15705v1 Announce Type: new Abstract: Reinforcement Fine-Tuning (RFT) has established itself as a critical paradigm for the alignment of Multi-modal Large Language Models (MLLMs) with comple

safetyarxiv-cs-lg
20 Apr 2026
Tools

🔬 Training Transformers to solve 95% failure rate of Cancer Trials — Ron Alfa & Daniel Bear, Noetik

DGX agent

Noetik uses transformer neural networks and machine learning to improve cancer drug trial success rates, addressing the historically high failure rate (95%) in clinical development. The approach lever

toolslatent-space
20 Apr 2026
Model Releases

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions

DGX agent

arXiv:2604.15823v1 Announce Type: new Abstract: Embodied robotic agents often perceive movies through an egocentric screen-view interface rather than native cinematic footage, introducing domain shift

model-releasesarxiv-cs-cv
20 Apr 2026
Agents

Weak-Link Optimization for Multi-Agent Reasoning and Collaboration

DGX agent

arXiv:2604.15972v1 Announce Type: new Abstract: LLM-driven multi-agent frameworks address complex reasoning tasks through multi-role collaboration. However, existing approaches often suffer from reaso

agentsarxiv-cs-ai
20 Apr 2026
Industry

We're opening a Hugging Face office in Tokyo! Our goal: help open-source AI develop in Japan and grow the local community. Let's meet! ハギングフ…

DGX agent

We're opening a Hugging Face office in Tokyo! Our goal: help open-source AI develop in Japan and grow the local community. Let's meet! ハギングフェイスの東京オフィスがオープンしました! 私たちの目標は、日本におけるオープンソースAIの発展を支援し、ローカルコミュニ

industryclem-delangue--x
20 Apr 2026
Local Ai

Where Do Vision-Language Models Fail? World Scale Analysis for Image Geolocalization

DGX agent

arXiv:2604.16248v1 Announce Type: new Abstract: Image geolocalization has traditionally been addressed through retrieval-based place recognition or geometry-based visual localization pipelines. Recent

local-aiarxiv-cs-cv
20 Apr 2026
Safety

Whose Facts Win? LLM Source Preferences under Knowledge Conflicts

DGX agent

arXiv:2601.03746v3 Announce Type: replace Abstract: As large language models (LLMs) are more frequently used in retrieval-augmented generation pipelines, it is increasingly relevant to study their beh

safetyarxiv-cs-cl
20 Apr 2026
Model Releases

Wisdom is Knowing What not to Say: Hallucination-Free LLMs Unlearning via Attention Shifting

DGX agent

arXiv:2510.17210v3 Announce Type: replace Abstract: The increase in computing power and the necessity of AI-assisted decision-making boost the growing application of large language models (LLMs). Alon

model-releasesarxiv-cs-cl
20 Apr 2026
Hardware

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed…

DGX agent

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed with building bigger LLMs. Trillions of parameters. Billion

hardwareyann-lecun--x
20 Apr 2026
Agents

Great paper on self-improving agents. Why? We need to think more deeply about AI agent system design. The protocol specifies a framework for…

DGX agent

Great paper on self-improving agents. Why? We need to think more deeply about AI agent system design. The protocol specifies a framework for proposing, assessing, and committing improvements with audi

agentsdair-ai--x
19 Apr 2026
Tutorials

in 1982 Titanic survivor Ruth Becker was giving an interview where she stated the ship broke in two. The treasurer of the Titanic Historical…

DGX agent

in 1982 Titanic survivor Ruth Becker was giving an interview where she stated the ship broke in two. The treasurer of the Titanic Historical Society actually took the microphone away from her and said

tutorialsjeremy-howard--x
19 Apr 2026
Model Releases

The continuing gap between the capabilities of Gemini Pro 3.1 (very good model) and the capabilities of the Gemini app/website is odd. The m…

DGX agent

The continuing gap between the capabilities of Gemini Pro 3.1 (very good model) and the capabilities of the Gemini app/website is odd. The model can do what Claude/GPT can do, but there is a minimal h

model-releasesethan-mollick--x
19 Apr 2026
Applications

Key to note that AI scientists are not experts on labor. Some other economists active on X doing work on AI & labor: @alexolegimas, @danielr…

DGX agent

Key to note that AI scientists are not experts on labor. Some other economists active on X doing work on AI & labor: @alexolegimas, @danielrock, @joshgans & @robseamans (among many others) But worth n

applicationsethan-mollick--x
18 Apr 2026
Model Releases

30K+ likes in the first hour. 🤯 That is crazy! Design is unsolved with agents. But lots of impactful work generated by agents is around des…

DGX agent

30K+ likes in the first hour. 🤯 That is crazy! Design is unsolved with agents. But lots of impactful work generated by agents is around design. Claude Design is Anthropic's way of saying that they are

model-releasesdair-ai--x
17 Apr 2026
Agents

AMA: Adaptive Memory via Multi-Agent Collaboration

DGX agent

arXiv:2601.20352v3 Announce Type: replace Abstract: The rapid evolution of Large Language Model (LLM) agents has necessitated robust memory systems to support cohesive long-term interaction and comple

agentsarxiv-cs-ai
17 Apr 2026
Applications

An Underexplored Frontier: Large Language Models for Rare Disease Patient Education and Communication -- A scoping review

DGX agent

arXiv:2604.14179v1 Announce Type: new Abstract: Rare diseases affect over 300 million people worldwide and are characterized by complex care pathways, limited clinical expertise, and substantial unmet

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

Assessment Design in the AI Era: A Method for Identifying Items Functioning Differentially for Humans and Chatbots

DGX agent

arXiv:2603.23682v2 Announce Type: replace-cross Abstract: The rapid adoption of large language models (LLMs) in education raises profound challenges for assessment design. To adapt assessments to the

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Benchmarking Optimizers for MLPs in Tabular Deep Learning

DGX agent

arXiv:2604.15297v1 Announce Type: new Abstract: MLP is a heavily used backbone in modern deep learning (DL) architectures for supervised learning on tabular data, and AdamW is the go-to optimizer used

model-releasesarxiv-cs-lg
17 Apr 2026
Model Releases

Correcting Suppressed Log-Probabilities in Language Models with Post-Transformer Adapters

DGX agent

arXiv:2604.14174v1 Announce Type: new Abstract: Alignment-tuned language models frequently suppress factual log-probabilities on politically sensitive topics despite retaining the knowledge in their h

model-releasesarxiv-cs-cl
17 Apr 2026
Industry

Counterpoint: India's smartphone shipments fell 3% YoY in Q1 2026, a six-year low, as price hikes weigh on sales; 80+ smartphone models saw price hikes of ~15% (Abhinav Parmar/Reuters)

DGX agent

Abhinav Parmar / Reuters: Counterpoint: India's smartphone shipments fell 3% YoY in Q1 2026, a six-year low, as price hikes weigh on sales; 80+ smartphone models saw price hikes of ~15% — India's smar

industrytechmeme
17 Apr 2026
Model Releases

Create Expert Content: Deploying a Multi-Agent System with Terraform and Cloud Run

DGX agent

In support of our mission to accelerate the developer journey on Google Cloud, we built Dev Signal: a multi-agent system designed to transform raw community signals into reliable technical guidance by

model-releasesgoogle-cloud-ai
17 Apr 2026
Tutorials

DETR-ViP: Detection Transformer with Robust Discriminative Visual Prompts

DGX agent

arXiv:2604.14684v1 Announce Type: new Abstract: Visual prompted object detection enables interactive and flexible definition of target categories, thereby facilitating open-vocabulary detection. Since

tutorialsarxiv-cs-cv
17 Apr 2026
Agents

DigiForest: Digital Analytics and Robotics for Sustainable Forestry

DGX agent

arXiv:2604.14652v1 Announce Type: new Abstract: Covering one third of Earth's land surface, forests are vital to global biodiversity, climate regulation, and human well-being. In Europe, forests and w

agentsarxiv-cs-ro
17 Apr 2026
Agents

Dual Pose-Graph Semantic Localization for Vision-Based Autonomous Drone Racing

DGX agent

arXiv:2604.15168v1 Announce Type: new Abstract: Autonomous drone racing demands robust real-time localization under extreme conditions: high-speed flight, aggressive maneuvers, and payload-constrained

agentsarxiv-cs-ro
17 Apr 2026
Model Releases

ECM Contracts: Contract-Aware, Versioned, and Governable Capability Interfaces for Embodied Agents

DGX agent

arXiv:2604.13097v1 Announce Type: cross Abstract: Embodied agents increasingly rely on modular capabilities that can be installed, upgraded, composed, and governed at runtime. Prior work has introduce

model-releasesarxiv-cs-ai
17 Apr 2026
Agents

Empowerment Gain and Causal Model Construction: Children and adults are sensitive to controllability and variability in their causal interventions

DGX agent

arXiv:2512.08230v2 Announce Type: replace Abstract: Learning about the causal structure of the world is a fundamental problem for human cognition. Causal models and especially causal learning have pro

agentsarxiv-cs-ai
17 Apr 2026
Agents

Enabling Agents to Communicate Entirely in Latent Space

DGX agent

arXiv:2511.09149v4 Announce Type: replace Abstract: While natural language is the de facto communication medium for LLM-based agents, it presents a fundamental constraint. The process of downsampling

agentsarxiv-cs-lg
17 Apr 2026
Model Releases

FoodSense: A Multisensory Food Dataset and Benchmark for Predicting Taste, Smell, Texture, and Sound from Images

DGX agent

arXiv:2604.14388v1 Announce Type: new Abstract: Humans routinely infer taste, smell, texture, and even sound from food images a phenomenon well studied in cognitive science. However, prior vision lang

model-releasesarxiv-cs-cv
17 Apr 2026
Applications

Grok 4.3 Beta is insanely good and now comes with powerful new tools that are super helpful for everyday use: • Create clean slides, graphs,…

DGX agent

Grok 4.3 Beta is insanely good and now comes with powerful new tools that are super helpful for everyday use: • Create clean slides, graphs, images and dashboards • Pull real-time data and images dire

applicationselon-musk--x
17 Apr 2026
Agents

@grok you go first

DGX agent

This post likely discusses Grok, an AI assistant developed by xAI, possibly exploring its capabilities, performance, or a specific interaction with the system. Given the casual phrasing and that it's

agentsyohei-nakajima--x
17 Apr 2026
← Previous
1…543544545546547…553
Next →