AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
Human
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,059 results
5 Jun 2026

CHALIS: A Challenge Dataset for Language Identification in Difficult Scenarios

Model ReleasesDGX agent

arXiv:2606.06088v1 Announce Type: new Abstract: We present CHALIS (Challenging Language Identification Samples), a new benchmark dataset explicitly designed to address difficult cases in language iden

Channel-Wise Mixed-Precision Quantization for Large Language Models

Model ReleasesDGX agent

arXiv:2410.13056v4 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable success across a wide range of language tasks, but their deployment on edge devices remain

ChartAttack: Testing the Vulnerability of LLMs to Malicious Prompting in Chart Generation

ResearchDGX agent

arXiv:2601.12983v3 Announce Type: replace Abstract: Multimodal large language models (MLLMs) are increasingly used to automate chart generation from data tables, improving analysis and reporting effic

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CHASE: Adversarial Red-Blue Teaming for Improving LLM Safety using Reinforcement Learning

SafetyDGX agent

arXiv:2606.05523v1 Announce Type: new Abstract: Despite advances in safety alignment, prompt-rewriting attacks such as persona modulation, fictional framing and persuasion-based reformulation, can byp

chat is he cooked

ToolsDGX agent

This post likely discusses whether someone or something is 'cooked' (slang for finished, defeated, or in trouble), though without access to the specific tweet content, the exact subject matter cannot

ChatGPT’s biggest memory upgrade starts rolling out.

IndustryDGX agent

OpenAI announced a new memory architecture built on its 'dreaming' background process that improves ChatGPT's ability to carry forward useful context, follow user preferences, and remain accurate as t

Check out the NVIDIA blog!

HardwareDGX agent

Check out the NVIDIA blog! This week at #CVPR2026, NVIDIA Research is presenting three papers across physical ai that offer groundbreaking solutions for training at scale across diverse applications:

'Chi nas dal soch el sent de legn' -- Auditing Text Corpora for Lombard

SafetyDGX agent

arXiv:2606.06349v1 Announce Type: new Abstract: Several of the world's languages are still under-resourced in terms of Natural Language Processing (NLP) tools. This is mostly due to the lack of high-q

China's securities regulator tightened oversight of the country's ~$3.4T private fund industry, and said it would encourage tech-focused VC investments (Reuters)

ApplicationsDGX agent

Reuters: China's securities regulator tightened oversight of the country's ~3.4T private fund industry, and said it would encourage tech-focused VC investments — China on Friday tightened oversight of

CLASH: Evaluating Language Models on Judging High-Stakes Dilemmas from Multiple Perspectives

Model ReleasesDGX agent

arXiv:2504.10823v4 Announce Type: replace Abstract: Navigating dilemmas involving conflicting values is challenging even for humans in high-stakes domains, let alone for AI, yet prior work has been li

CLEAR: Cognition and Latent Evaluation for Adaptive Routing in End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.06219v1 Announce Type: new Abstract: End-to-end autonomous driving models often struggle to balance multi-modal maneuver generation with real-time inference constraints. While diffusion mod

CLFEC: A New Task for Unified Linguistic and Factual Error Correction in paragraph-level Chinese Professional Writing

Model ReleasesDGX agent

arXiv:2602.23845v2 Announce Type: replace Abstract: Chinese text correction has traditionally focused on spelling and grammar, while factual error correction is usually treated separately. However, in

Code2LoRA: Hypernetwork-Generated Adapters for Code Language Models under Software Evolution

Model ReleasesDGX agent

arXiv:2606.06492v1 Announce Type: cross Abstract: Code language models need repository-level context to resolve imports, APIs, and project conventions. Existing methods inject this knowledge as long i

Coding with 'Enemy': Can Human Developers Detect AI Agent Sabotage?

Model ReleasesDGX agent

arXiv:2606.05647v1 Announce Type: cross Abstract: AI coding agents are increasingly embedded in real-world software development, collaborating with human developers while gaining broader access to cod

CoFi-UCGen: Coarse-to-Fine Unsupervised Conditional Generation without Label Priors

TutorialsDGX agent

arXiv:2606.05652v1 Announce Type: new Abstract: Unsupervised conditional image generation (UCGen) aims to control generation without relying on manually annotated labels, yet remains challenging due t

ColBERTSaR: Sparsified ColBERT Index via Product Quantization

ResearchDGX agent

arXiv:2606.05568v1 Announce Type: cross Abstract: While ColBERT is an effective neural retrieval architecture, it requires a heavy index structure to support candidate set retrieval based on approxima

CollabBench: Benchmarking and Unleashing Collaborative Ability of LLMs with Diverse Players via Proactive Engagement

Model ReleasesDGX agent

arXiv:2606.05793v1 Announce Type: new Abstract: While LLM-based agents excel at individual tasks, effective collaboration with realistic human partners remains challenging. Most of the existing conver

CollabSim: A CSCW-Grounded Methodology for Investigating Collaborative Competence of LLM Agents through Controlled Multi-Agent Experiments

AgentsDGX agent

arXiv:2606.06399v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models have shown growing promise, with their effectiveness resting on agents' ability to coordinate t

Come visit our CVPR booth for hands-on experiences: – Neural video generation + sim drive – Selfies with @Tesla_Optimus – Real FSD test driv…

AgentsDGX agent

Come visit our CVPR booth for hands-on experiences: – Neural video generation + sim drive – Selfies with @Tesla_Optimus – Real FSD test drives in Tesla models of your choice Chat directly with the tea

CoMoL: Efficient Mixture of LoRA Experts via Dynamic Core Space Merging

Model ReleasesDGX agent

arXiv:2603.00573v2 Announce Type: replace Abstract: Large language models (LLMs) achieve remarkable performance on diverse downstream and domain-specific tasks via parameter-efficient fine-tuning (PEF

Comparison of Deep Learning Frameworks For Rice Disease Mapping From UAV Multispectral Imaging

ResearchDGX agent

arXiv:2606.06359v1 Announce Type: new Abstract: In this study, UAV multispectral imagery is used to segment the severity of bacterial leaf blight (BLB) in rice using convolutional neural networks (CNN

Complexity-Balanced Diffusion Splitting

Local AiDGX agent

arXiv:2606.06477v1 Announce Type: new Abstract: Standard continuous-time generative models rely on monolithic architectures that must navigate vastly different signal regimes, from isotropic noise to

ComplexityMT: Benchmarking the Interaction Between Text Complexity and Machine Translation

ResearchDGX agent

arXiv:2606.05421v1 Announce Type: new Abstract: When a text is translated, does the translation retain the complexity of the original? We introduce ComplexityMT, a new challenge for assessing how text

Compress-Distill: Reasoning Trace Compression for Efficient Knowledge Distillation

Model ReleasesDGX agent

arXiv:2606.05988v1 Announce Type: cross Abstract: Reasoning models produce long chain-of-thought traces that are costly to distill and encourage verbose student outputs. We study post-hoc compression

Computation-Aware Event-to-Frame Reconstruction via Selective Attention

ResearchDGX agent

arXiv:2606.06142v1 Announce Type: new Abstract: Event-to-frame (E2F) reconstruction bridges asynchronous event streams with frame-based vision pipelines, but existing methods often face a trade-off be

Congrats to Reardon and team on @flourishailabs. If they can get AI sample efficiency and energy consumption to human levels, thats going to…

ResearchDGX agent

Congrats to Reardon and team on @flourishailabs. If they can get AI sample efficiency and energy consumption to human levels, thats going to change so many things in the world! And we are live! https:

Contextualized Prompting For Stance Detection On Social Media

Model ReleasesDGX agent

arXiv:2606.06022v1 Announce Type: new Abstract: Stance detection on social media is challenging due to short, noisy, and context-dependent language. While large language models (LLMs) show zero-shot g

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments

Model ReleasesDGX agent

arXiv:2606.05661v1 Announce Type: cross Abstract: Continual learning, the ability of AI systems to improve through sequential experience, has attracted substantial interest, but no high-quality benchm

Correcting Prompt Dependence in LLM Benchmarks: A Bayesian Hierarchical Model with Embedding-Space Clustering

ResearchDGX agent

arXiv:2510.05709v2 Announce Type: replace-cross Abstract: LLM benchmarking metrics often misstate performance and uncertainty as they rely on two assumptions that frequently do not hold in practice: (

Corruption

IndustryDGX agent

Corruption Let me get this straight > 84% of Americans support requiring a photo id to vote > every major democracy in the world does this except the US > even under our very eyes, the mail-in ballots

Cosine Misleads: Auxiliary Losses Reshape Vision Language Models, Not Their Latents

SafetyDGX agent

arXiv:2606.05753v1 Announce Type: new Abstract: Latent visual reasoning (LVR) inserts supervised latent tokens between perception and answer generation in vision-language models (VLMs). The field uses

CoT-Space: A Theoretical Framework for Internal Slow-Thinking via Reinforcement Learning

ResearchDGX agent

arXiv:2509.04027v3 Announce Type: replace-cross Abstract: Test-time scaling, primarily manifested through multi-step Chain-of-Thought (CoT) reasoning via Reinforcement Learning (RL), has emerged as a

Cowork is at its best on work that’s too big for a chat: research across dozens of accounts, recurring reports, triaging my inbox and drafti…

ToolsDGX agent

Cowork is at its best on work that’s too big for a chat: research across dozens of accounts, recurring reports, triaging my inbox and drafting replies. If you’ve been curious, this is a good month to

Critical context on the new Anthropic blog: 1, AGI is *harder* than RSI (as used below). AGI: machine can do anything human can do, autonomo…

Model ReleasesDGX agent

Critical context on the new Anthropic blog: 1, AGI is *harder* than RSI (as used below). AGI: machine can do anything human can do, autonomously [not achieved] RSI (as used below): AI is a useful codi

Data + AI Summit 2026: Insider’s Guide for Financial Services Leaders

TutorialsDGX agent

This Databricks guide provides financial services leaders with essential information and strategies for attending or engaging with the Data + AI Summit 2026, covering industry-specific insights on lev

Decomposing Factual Sycophancy in Language Models: How Size and Instruction Tuning Shape Robustness

ResearchDGX agent

arXiv:2606.06306v1 Announce Type: new Abstract: Factual sycophancy occurs when a language model abandons a correct, verifiable answer under social pressure. Because a flip occurs only when pressure to

Deep Learning-assisted AMD Staging based on OCT and OCT Angiography

ResearchDGX agent

arXiv:2606.05379v1 Announce Type: new Abstract: To develop and evaluate deep learning models for automated grading of age-related macular degeneration (AMD) severity using optical coherence tomography

Deep Learning-based 3D Oral Cavity Reconstruction Using 2D Intraoral Images

ResearchDGX agent

arXiv:2606.05998v1 Announce Type: new Abstract: Oral 3D modelling is one of the most essential stages in dentistry, and many different approaches, such as impression taking and intraoral scanning, are

Dense Contexts Are Hard Contexts: Lexical Density Limits Effective Context in LLMs

Model ReleasesDGX agent

arXiv:2606.06203v1 Announce Type: new Abstract: Input length and the position of relevant information are widely cited as the primary causes of degraded LLM long-context performance. Here, we study le

DexFuture: Hierarchical Future-State Visuomotor Targeting for Bimanual Dexterous Tool Use

SafetyDGX agent

arXiv:2606.05699v1 Announce Type: new Abstract: Bimanual dexterous tool use remains challenging for robots due to high-dimensional hand configurations and complex hand-tool-object dynamics and contact

Diff-CA: Separating Common and Salient Factors with Diffusion Models

TutorialsDGX agent

arXiv:2606.06120v1 Announce Type: new Abstract: Contrastive Analysis aims to separate factors that are common between two data distributions from those that are salient to only one of them. Existing c

DiG-Plan: Mitigating Early Commitment for Tool-Graph Planning via Diffusion Guidance

ResearchDGX agent

arXiv:2606.05728v1 Announce Type: cross Abstract: Generating executable tool plans requires selecting appropriate subsets from tool libraries, a combinatorial search problem with an exponentially larg

DisasterBench: A Multimodal Benchmark for UAV-Based Disaster Response in Complex Environments

Model ReleasesDGX agent

arXiv:2606.06217v1 Announce Type: new Abstract: When a disaster unfolds, responders must answer not only what is happening, but also why it is happening, what will happen next, and what to do now, oft

Discrete-WAM: Unified Discrete Vision-Action Token Editing for World-Policy Learning

SafetyDGX agent

arXiv:2606.05645v1 Announce Type: new Abstract: Autonomous driving requires reasoning about how ego actions shape the evolution of the surrounding world. However, most end-to-end methods rely on direc

Disentangled Fine-Grained Prototype Learning for Incomplete Image-Tabular Classification

SafetyDGX agent

arXiv:2606.05455v1 Announce Type: new Abstract: The missing-modality problem poses a significant challenge in image-tabular multimodal learning across a wide range of multimedia applications, includin

Do MLLMs Capture How Interfaces Guide User Behavior? A Benchmark for Multimodal UI/UX Design Understanding

Model ReleasesDGX agent

arXiv:2505.05026v5 Announce Type: replace Abstract: User interface (UI) design goes beyond visuals to shape user experience (UX), underscoring the shift toward UI/UX as a unified concept. While recent

Do Models Share Safety Representations? Cross-Model Steering for Safe Visual Generation

SafetyDGX agent

arXiv:2606.05290v1 Announce Type: new Abstract: Recent progress in generative modeling has made safety control a central challenge, yet existing approaches remain largely model-specific, requiring ret

DocHop-QA: Towards Multi-Hop Reasoning over Multimodal Document Collections

Model ReleasesDGX agent

arXiv:2508.15851v2 Announce Type: replace Abstract: Despite rapid progress in large language models (LLMs), current QA benchmarks still overlook the core challenge of real-world scientific information

Domain-Aware Mispronunciation Detection and Diagnosis Using Language-Specific Statistical Graphs

Model ReleasesDGX agent

arXiv:2606.05569v1 Announce Type: new Abstract: Mispronunciation Detection and Diagnosis (MDD) has gained increasing importance in computer-assisted language learning and speech technology in recent y

Domain-Conditioned Safety in Frontier Computer-Using Agents: A 793-Episode Browser Benchmark, a Coding-Domain Cross-Reference, and a Reproducibility Audit of Recent Red-Teaming

Model ReleasesDGX agent

arXiv:2606.05233v1 Announce Type: cross Abstract: Recent computer-using-agent (CUA) red-teaming papers report prompt-injection attack success rates (ASR) of 42-98%, but these headline numbers cluster

DRIFT: A Residual Flow Adapter for Decoding Continuous Outputs in Vision-Language Models

ResearchDGX agent

arXiv:2606.05758v1 Announce Type: new Abstract: Many modern vision-language models (VLMs) build on autoregressive decoding of discrete tokens. While text-based output interfaces enable scalable pretra

Drishti AI-Event Guardian: An Intelligent Real-Time Crowd Monitoring and Emergency Response System for Mass Gathering Events

SafetyDGX agent

arXiv:2606.05185v1 Announce Type: cross Abstract: Mass gathering events are associated with critical safety incidents caused by insufficient crowd monitoring and inadequate emergency response coordina

Drive-KD: Multi-Teacher Distillation for VLMs in Autonomous Driving

Model ReleasesDGX agent

arXiv:2601.21288v2 Announce Type: replace-cross Abstract: Autonomous driving is an important and safety-critical task, and recent advances in LLMs/VLMs have opened new possibilities for reasoning and

Drives for Vercel Sandbox in Private Beta

ToolsDGX agent

Vercel announced a new feature called 'Drives' for Vercel Sandbox, which entered private beta testing. This feature likely provides persistent storage capabilities or file management functionality wit

Dual Feature Decoupling for Fine-Grained OOD Detection

ApplicationsDGX agent

arXiv:2606.05536v1 Announce Type: new Abstract: Out-of-distribution detection (OOD) is an indispensable technique when applying machine learning models to real-world scenarios. Most existing OOD detec

Dynamic Multi-Agent Pickup and Delivery in Robotic Cellular Warehousing Systems

AgentsDGX agent

arXiv:2606.05669v1 Announce Type: new Abstract: Robotic Cellular Warehousing Systems (RCWS) give rise to multi-agent pickup and delivery (MAPD) processes in which robots sequentially collect multiple

Dynamic Thinking-Token Selection for Efficient Reasoning in Large Reasoning Models

ResearchDGX agent

arXiv:2601.18383v2 Announce Type: replace-cross Abstract: Large Reasoning Models (LRMs) excel at solving complex problems by explicitly generating a reasoning trace before deriving the final answer. H

EasyLens: A Training-Free Plug-and-Play Subtle-Lesion Representation Amplifier for Medical Vision-Language Models

Local AiDGX agent

arXiv:2606.06379v1 Announce Type: new Abstract: Medical vision-language models (VLMs) have shown increasing potential for clinical image interpretation, including lesion detection and report generatio

EDIT: Evidence-Diagnosed Intervention Training for Rule-Faithful LLM Grading

ApplicationsDGX agent

arXiv:2606.06350v1 Announce Type: new Abstract: Reliable rubric grading requires more than accurate score prediction. Each judgement must be grounded in the mark scheme and evidence from the student a

Efficient Computation of Distance Functions for Navigation Vector Fields in Lie Groups

ResearchDGX agent

arXiv:2606.05372v1 Announce Type: new Abstract: Vector-field-based methods are widely used for robot control and are often applied to the path-tracking problem. Some vector field approaches require re

← Previous
1…683684685686687…1518
Next →