AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,794 results
Safety

Personality Shapes Gender Bias in Persona-Conditioned LLM Narratives Across English and Hindi: An Empirical Investigation

DGX agent

arXiv:2604.23600v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed in persona-driven applications such as education, customer service, and social platforms, where m

safetyarxiv-cs-cl
28 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Position: Logical Soundness is not a Reliable Criterion for Neurosymbolic Fact-Checking with LLMs

DGX agent

arXiv:2604.04177v2 Announce Type: replace Abstract: As large language models (LLMs) are increasing integrated into fact-checking pipelines, formal logic is often proposed as a rigorous means by which

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

Pref-CTRL: Preference Driven LLM Alignment using Representation Editing

DGX agent

arXiv:2604.23543v1 Announce Type: cross Abstract: Test-time alignment methods offer a promising alternative to fine-tuning by steering the outputs of large language models (LLMs) at inference time wit

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Process Supervision of Confidence Margin for Calibrated LLM Reasoning

DGX agent

arXiv:2604.23333v1 Announce Type: cross Abstract: Scaling test-time computation with reinforcement learning (RL) has emerged as a reliable path to improve large language models (LLM) reasoning ability

researcharxiv-cs-cl
28 Apr 2026
Tutorials

Propagation Structure-Semantic Transfer Learning for Robust Fake News Detection

DGX agent

arXiv:2604.23974v1 Announce Type: new Abstract: Fake news generally refers to false information that is spread deliberately to deceive people, which has detrimental social effects. Existing fake news

tutorialsarxiv-cs-cl
28 Apr 2026
Model Releases

Reading in the Dark: Low-light Scene Text Recognition

DGX agent

arXiv:2604.23685v1 Announce Type: new Abstract: Accurate text recognition in low-light environments is essential for intelligent systems in applications ranging from autonomous vehicles to smart surve

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

Rethinking Parameter Sharing for LLM Fine-Tuning with Multiple LoRAs

DGX agent

arXiv:2509.25414v2 Announce Type: replace-cross Abstract: Large language models are often adapted using parameter-efficient techniques such as Low-Rank Adaptation (LoRA), formulated as y = W_0x + BAx,

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Scalable LLM-based Coding of Dialogue in Healthcare Simulation: Balancing Coding Performance, Processing Time, and Environmental Impact

DGX agent

arXiv:2604.23255v1 Announce Type: cross Abstract: Research shows that dialogue, the interactive process through which participants articulate their thinking, plays a central role in constructing share

applicationsarxiv-cs-ai
28 Apr 2026
Model Releases

SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation

DGX agent

arXiv:2604.23274v1 Announce Type: new Abstract: Semi-supervised learning addresses label scarcity and high annotation costs in medical image segmentation by exploiting the latent information in unlabe

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

DGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

model-releasesarxiv-cs-cl
28 Apr 2026
Research

StereoFoley: Object-Aware Stereo Audio Generation from Video

DGX agent

We present StereoFoley, a video-to-audio generation framework that produces semantically aligned, temporally synchronized, and spatially accurate stereo sound at 48 kHz. While recent generative video-

researchapple-ml-research
28 Apr 2026
Model Releases

StructAlign: Structured Cross-Modal Alignment for Continual Text-to-Video Retrieval

DGX agent

arXiv:2601.20597v2 Announce Type: replace Abstract: Continual Text-to-Video Retrieval (CTVR) is a challenging multimodal continual learning setting, where models must incrementally learn new semantic

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SwissGov-RSD: A Human-annotated, Cross-lingual Benchmark for Token-level Recognition of Semantic Differences Between Related Documents

DGX agent

arXiv:2512.07538v3 Announce Type: replace Abstract: Recognizing semantic differences across documents is crucial for text generation evaluation and content alignment, especially in cross-lingual setti

model-releasesarxiv-cs-cl
28 Apr 2026
Research

Symbolic recovery of PDEs from measurement data

DGX agent

arXiv:2602.15603v2 Announce Type: replace Abstract: Models based on partial differential equations (PDEs) are powerful for describing a wide range of complex phenomena in the natural sciences. Accurat

researcharxiv-cs-lg
28 Apr 2026
Research

Task-guided Spatiotemporal Network with Diffusion Augmentation for EEG-based Dementia Diagnosis and MMSE Prediction

DGX agent

arXiv:2604.23964v1 Announce Type: cross Abstract: Patients with dementia typically exhibit cognitive impairment, which is routinely assessed using the Mini-Mental State Examination (MMSE). Concurrentl

researcharxiv-cs-ai
28 Apr 2026
Safety

The Consensus Trap: Dissecting Subjectivity and the 'Ground Truth' Illusion in Data Annotation

DGX agent

arXiv:2602.11318v3 Announce Type: replace Abstract: In machine learning, 'ground truth' refers to the assumed correct labels used to train and evaluate models. However, the foundational 'ground truth'

safetyarxiv-cs-ai
28 Apr 2026
Agents

Think Anywhere in Code Generation

DGX agent

arXiv:2603.29957v3 Announce Type: replace-cross Abstract: Recent advances in reasoning Large Language Models (LLMs) have primarily relied on upfront thinking, where reasoning occurs before final answe

agentsarxiv-cs-lg
28 Apr 2026
Local Ai

Unrealized Expectations: Comparing AI Methods vs Classical Algorithms for Maximum Independent Set

DGX agent

arXiv:2502.03669v3 Announce Type: replace-cross Abstract: AI methods, such as generative models and reinforcement learning, have recently been applied to combinatorial optimization (CO) problems, espe

local-aiarxiv-cs-ai
28 Apr 2026
Agents

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work

DGX agent

arXiv:2604.23674v1 Announce Type: new Abstract: With the emergence of large language models (LLMs) and AI agent frameworks, the human-AI co-work paradigm known as Vibe Coding is changing how people co

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

WebSerial Vision Training for Microcontrollers: A Browser-Based Companion to On-Device CNN Training

DGX agent

arXiv:2604.22834v1 Announce Type: new Abstract: This paper presents webmcu-vision-web, a single-file, zero-install browser application for end-to-end TinyML vision model training and deployment on the

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

What Did They Mean? How LLMs Resolve Ambiguous Social Situations across Perspectives and Roles

DGX agent

arXiv:2604.23942v1 Announce Type: cross Abstract: People increasingly turn to large language models (LLMs) to interpret ambiguous social situations: a delayed text reply, an unusually cold supervisor,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

When Corrective Hints Hurt: Prompt Design in Reasoner-Guided Repair of LLM Overcaution on Entailed Negations under OWL~2~DL

DGX agent

arXiv:2604.23398v1 Announce Type: new Abstract: We report a reproducible error pattern in GPT-5.4 on OWL~2~DL compliance queries: the model frequently answers ``unknown'' when the reasoner-entailed an

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Adapting MLLMs for Nuanced Video Retrieval

DGX agent

arXiv:2512.13511v2 Announce Type: replace Abstract: Our objective is to build an embedding model that captures the nuanced relationship between a search query and candidate videos. We cover three aspe

researcharxiv-cs-cv
27 Apr 2026
Safety

AgentBound: Securing Execution Boundaries of AI Agents

DGX agent

arXiv:2510.21236v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have evolved into AI agents that interact with external tools and environments to perform complex tasks. The Mode

safetyarxiv-cs-ai
27 Apr 2026
Model Releases

Anatomy-Aware Unsupervised Detection and Localization of Retinal Abnormalities in Optical Coherence Tomography

DGX agent

arXiv:2604.22139v1 Announce Type: new Abstract: Reliable automated analysis of Optical Coherence Tomography (OCT) imaging is crucial for diagnosing retinal disorders but faces a critical barrier: the

model-releasesarxiv-cs-cv
27 Apr 2026
Research

Chain-of-Memory: Lightweight Memory Construction with Dynamic Evolution for LLM Agents

DGX agent

arXiv:2601.14287v2 Announce Type: replace Abstract: External memory systems are pivotal for enabling Large Language Model (LLM) agents to maintain persistent knowledge and perform long-horizon decisio

researcharxiv-cs-lg
27 Apr 2026
Tutorials

CLVAE: A Variational Autoencoder for Long-Term Customer Revenue Forecasting

DGX agent

arXiv:2604.22636v1 Announce Type: cross Abstract: Predicting customers' long-term revenue from sparse and irregular transaction data is central to marketing resource allocation in non-contractual sett

tutorialsarxiv-cs-lg
27 Apr 2026
Safety

Controllable Spoken Dialogue Generation: An LLM-Driven Grading System for K-12 Non-Native English Learners

DGX agent

arXiv:2604.22542v1 Announce Type: cross Abstract: Large language models (LLMs) often fail to meet the pedagogical needs of K-12 English learners in non-native contexts due to a proficiency mismatch. T

safetyarxiv-cs-ai
27 Apr 2026
Research

Decomposed Attention Fusion in MLLMs for Training-Free Video Reasoning Segmentation

DGX agent

arXiv:2510.19592v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) demonstrate strong video understanding by attending to visual tokens relevant to textual queries. To direct

researcharxiv-cs-cv
27 Apr 2026
Model Releases

EV-CLIP: Efficient Visual Prompt Adaptation for CLIP in Few-shot Action Recognition under Visual Challenges

DGX agent

arXiv:2604.22595v1 Announce Type: new Abstract: CLIP has demonstrated strong generalization in visual domains through natural language supervision, even for video action recognition. However, most exi

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

@FireworksAI_HQ deeply believes in delivering the frontier quality. We will spend all the effort giving our users and the broad community th…

DGX agent

@FireworksAI_HQ deeply believes in delivering the frontier quality. We will spend all the effort giving our users and the broad community the best OSS model quality. Deepseek V4 API up anytime now aft

model-releasesfireworks-ai--x
27 Apr 2026
Agents

For the past few years, humans have been doing “prompt engineering” to coax the best performance out of different LLMs. In this work, we exp…

DGX agent

For the past few years, humans have been doing “prompt engineering” to coax the best performance out of different LLMs. In this work, we explored what happens if we train an AI to do that job instead.

agentsdavid-ha--x
27 Apr 2026
Safety

How Many Visual Levers Drive Urban Perception? Interventional Counterfactuals via Multiple Localised Edits

DGX agent

arXiv:2604.22103v1 Announce Type: cross Abstract: Street-view perception models predict subjective attributes such as safety at scale, but remain correlational: they do not identify which localized vi

safetyarxiv-cs-cv
27 Apr 2026
Model Releases

how to adjust the thinking effort for deepseek v4 on ollama cloud

DGX agent

DeepSeek V4 models on Ollama Cloud support three thinking modes: 'No thinking' for fast answers, 'Thinking' for careful analysis, and 'Max thinking' for maximum reasoning effort . Users can adjust thi

model-releasesr-ollama
27 Apr 2026
Research

LayerBoost: Layer-Aware Attention Reduction for Efficient LLMs

DGX agent

arXiv:2604.22050v1 Announce Type: cross Abstract: Transformers are mostly relying on softmax attention, which introduces quadratic complexity with respect to sequence length and remains a major bottle

researcharxiv-cs-cl
27 Apr 2026
Model Releases

LLMs as Assessors: Right for the Right Reason?

DGX agent

arXiv:2601.08919v2 Announce Type: replace-cross Abstract: A good deal of recent research has focused on how Large Language Models (LLMs) may be used as judges in place of humans to evaluate the qualit

model-releasesarxiv-cs-cl
27 Apr 2026
Research

NiuTrans.LMT: Toward Inclusive and Scalable Multilingual Machine Translation with LLMs

DGX agent

arXiv:2511.07003v2 Announce Type: replace Abstract: Large language models have significantly advanced Multilingual Machine Translation (MMT), yet scaling to many languages while keeping quality robust

researcharxiv-cs-cl
27 Apr 2026
Model Releases

OccDirector: Language-Guided Behavior and Interaction Generation in 4D Occupancy Space

DGX agent

arXiv:2604.22240v1 Announce Type: new Abstract: Generative world models increasingly rely on 4D occupancy for realistic autonomous driving simulation. However, existing generation frameworks depend on

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

PL-MTEB: Polish Massive Text Embedding Benchmark

DGX agent

arXiv:2405.10138v2 Announce Type: replace Abstract: In this paper, we introduce the Polish Massive Text Embedding Benchmark (PL-MTEB), a comprehensive benchmark for text embeddings in the Polish langu

model-releasesarxiv-cs-cl
27 Apr 2026
Research

PreMoE: Proactive Inference for Efficient Mixture-of-Experts

DGX agent

arXiv:2505.17639v3 Announce Type: replace Abstract: Mixture-of-Experts (MoE) models offer dynamic computation, but are typically deployed as static full-capacity models, missing opportunities for depl

researcharxiv-cs-lg
27 Apr 2026
Research

PrivUn: Unveiling Latent Ripple Effects and Shallow Forgetting in Privacy Unlearning

DGX agent

arXiv:2604.22076v1 Announce Type: cross Abstract: Large language models (LLMs) often memorize private information during training, raising serious privacy concerns. While machine unlearning has emerge

researcharxiv-cs-cl
27 Apr 2026
Model Releases

Region Matters: Efficient and Reliable Region-Aware Visual Place Recognition

DGX agent

arXiv:2604.22390v1 Announce Type: new Abstract: Visual Place Recognition (VPR) determines a query image's geographic location by matching it against geotagged databases. However, existing methods stru

model-releasesarxiv-cs-cv
27 Apr 2026
Model Releases

ResRank: Unifying Retrieval and Listwise Reranking via End-to-End Joint Training with Residual Passage Compression

DGX agent

arXiv:2604.22180v1 Announce Type: cross Abstract: Large language model (LLM) based listwise reranking has emerged as the dominant paradigm for achieving state-of-the-art ranking effectiveness in infor

model-releasesarxiv-cs-ai
27 Apr 2026
Research

Rethinking Math Reasoning Evaluation: A Robust LLM-as-a-Judge Framework Beyond Symbolic Rigidity

DGX agent

arXiv:2604.22597v1 Announce Type: new Abstract: Recent advancements in large language models have led to significant improvements across various tasks, including mathematical reasoning, which is used

researcharxiv-cs-ai
27 Apr 2026
Applications

Segment Any-Quality Images with Generative Latent Space Enhancement

DGX agent

arXiv:2503.12507v3 Announce Type: replace Abstract: Despite their success, Segment Anything Models (SAMs) experience significant performance drops on severely degraded, low-quality images, limiting th

applicationsarxiv-cs-cv
27 Apr 2026
Research

SSG: Logit-Balanced Vocabulary Partitioning for LLM Watermarking

DGX agent

arXiv:2604.22438v1 Announce Type: cross Abstract: Watermarking has emerged as a promising technique for tracing the authorship of content generated by large language models (LLMs). Among existing appr

researcharxiv-cs-ai
27 Apr 2026
Industry

This is how it's done! Who else should we ask to release weights?

DGX agent

Clem Delangue, CEO of Hugging Face, discusses best practices for releasing model weights in the open-source AI community and advocates for other AI organizations to follow suit in making their models

industryclem-delangue--x
27 Apr 2026
Safety

Toward Principled LLM Safety Testing: Solving the Jailbreak Oracle Problem

DGX agent

arXiv:2506.17299v2 Announce Type: replace-cross Abstract: As large language models (LLMs) become increasingly deployed in safety-critical applications, the lack of systematic methods to assess their v

safetyarxiv-cs-ai
27 Apr 2026
← Previous
1…549550551552553…1371
Next →