AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,759 results
Model Releases

WhoSaidIt: Human-LLM Collaborative Annotation for Text-Based Multilingual Speaker-Attribute Classification

DGX agent

arXiv:2605.26070v1 Announce Type: new Abstract: Annotating speaker attributes from text is inherently ambiguous, particularly in multilingual settings where demographic and social cues are implicit an

model-releasesarxiv-cs-cl
26 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

WISE: Web Information Satire and Fakeness Evaluation

DGX agent

arXiv:2512.24000v3 Announce Type: replace Abstract: Distinguishing fake or untrue news from satire or humor poses a unique challenge due to their overlapping linguistic features and divergent intent.

applicationsarxiv-cs-cl
26 May 2026
Research

Word Class Representations Spontaneously Emerge from Successor Representations Trained on Natural Language

DGX agent

arXiv:2605.24585v1 Announce Type: new Abstract: Language models are typically trained to predict the next token in a sequence. Here, we explore an alternative predictive principle from reinforcement l

researcharxiv-cs-cl
26 May 2026
Model Releases

A Comparative Evaluation of Structural Topic Models and BERTopic for Short, Open-Ended Survey Responses

DGX agent

arXiv:2605.23093v1 Announce Type: new Abstract: Topic modeling in applied psychology increasingly spans two methodological traditions: probabilistic bag-of-words models and newer embedding-based appro

model-releasesarxiv-cs-cl
25 May 2026
Research

A graph-based analysis of semantic types and coercion in contextualized word embeddings

DGX agent

arXiv:2605.23710v1 Announce Type: new Abstract: Semantic type mismatch between a noun and its context is central to coercion phenomena. This paper introduces a graph-based method to examine how lexica

researcharxiv-cs-cl
25 May 2026
Model Releases

A Reproducible Universal Dependencies-Style Pipeline for Katharevousa Greek Parliamentary Text

DGX agent

arXiv:2605.22978v1 Announce Type: new Abstract: Katharevousa Greek remains poorly served by contemporary NLP pipelines despite its importance for legal, administrative, and parliamentary archives. We

model-releasesarxiv-cs-cl
25 May 2026
Research

A Survey of Text and Speech Resources for Hausa and Fongbe: Availability, Quality, and Gaps for NLP Development

DGX agent

arXiv:2605.22828v1 Announce Type: new Abstract: This survey provides a comprehensive catalog of publicly available text and speech resources for two West African languages: Hausa, an Afroasiatic langu

researcharxiv-cs-cl
25 May 2026
Research

AI-Friendly LaTeX: Using LaTeX Code as a Knowledge Source for Retrieval-Augmented Generation

DGX agent

arXiv:2605.22923v1 Announce Type: cross Abstract: Large language models can answer questions about textbooks, lecture notes, and programming exercises more reliably when their answers are grounded in

researcharxiv-cs-cl
25 May 2026
Model Releases

AraHopeCorpus: Annotation Guidelines and Dataset for Hope Speech in Arabic Social Media Crisis Discourse

DGX agent

arXiv:2605.23325v1 Announce Type: new Abstract: Social media has become a crucial arena for shaping public narratives during armed conflicts, providing space for both harmful and constructive communic

model-releasesarxiv-cs-cl
25 May 2026
Applications

ARES: Automated Rubric Synthesis for Scalable LLM Reinforcement Learning

DGX agent

arXiv:2605.23454v1 Announce Type: new Abstract: Rubric-based rewards offer a promising way to extend reinforcement learning (RL) for large language models beyond tasks with automatically verifiable an

applicationsarxiv-cs-cl
25 May 2026
Applications

Articulatory strategy as a source of variation in acoustic vowel dynamics

DGX agent

arXiv:2605.23416v1 Announce Type: new Abstract: Acoustic vowel dynamics have some speaker-identifying characteristics, which have been ascribed to individual properties of articulatory strategies: for

applicationsarxiv-cs-cl
25 May 2026
Model Releases

Asking For An Old Friend: Diagnosing and Mitigating Temporal Failure Modes in LLM-based Statutory Question Answering

DGX agent

arXiv:2605.23497v1 Announce Type: new Abstract: Large language models are increasingly used for legal research, yet their fixed training cutoffs and reliance on static parametric knowledge are at odds

model-releasesarxiv-cs-cl
25 May 2026
Research

Benchmarking Gaslighting Attacks Against Speech Large Language Models

DGX agent

arXiv:2509.19858v2 Announce Type: replace Abstract: As Speech Large Language Models (Speech LLMs) become increasingly integrated into voice-based applications, ensuring their robustness against manipu

researcharxiv-cs-cl
25 May 2026
Model Releases

Benchmarking Google Embeddings 2 against Open-Source Models for Multilingual Dense Retrieval and RAG Systems

DGX agent

arXiv:2605.23618v1 Announce Type: new Abstract: We benchmark Google Embeddings (GE2), a Vertex-AI-hosted bi-encoder with 2,048-token context and explicit task-type conditioning, against five open-sour

model-releasesarxiv-cs-cl
25 May 2026
Research

Beyond Log Likelihood: Probability-Based Objectives for Supervised Fine-Tuning across the Model Capability Continuum

DGX agent

arXiv:2510.00526v3 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is the standard approach for post-training large language models (LLMs), yet it often shows limited generalization. We

researcharxiv-cs-cl
25 May 2026
Model Releases

BURMESE-SAN: Burmese NLP Benchmark for Evaluating Large Language Models

DGX agent

arXiv:2602.18788v3 Announce Type: replace Abstract: We introduce BURMESE-SAN, the first holistic benchmark that systematically evaluates large language models (LLMs) for Burmese across three core NLP

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Can AI Guess What You Know? Performance Comparison of Large Language Models for Human Domain Knowledge Estimation From Communication Logs

DGX agent

arXiv:2605.22971v1 Announce Type: new Abstract: Employees often struggle to identify ``who knows what,'' leading to organizational productivity losses. We investigate whether Large Language Models (LL

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

ChartFI: Benchmarking Faithfulness and Insightfulness of Chart Descriptions from Multimodal Large Language Models

DGX agent

arXiv:2605.23694v1 Announce Type: new Abstract: Chart descriptions are essential for accessibility, cross-modal retrieval, and assisting readers in extracting insights from complex visualizations. As

model-releasesarxiv-cs-cl
25 May 2026
Safety

ClimateChat-300K: A Multi-Modal Facebook Dataset for Understanding Diverse Perspectives in Climate Communication

DGX agent

arXiv:2605.23326v1 Announce Type: new Abstract: We present ClimateChat-300K, a large-scale dataset of 299,329 public Facebook posts about climate change collected between May 2020 and May 2024 through

safetyarxiv-cs-cl
25 May 2026
Local Ai

CultivAgents: Cultivating Relationship-Centered Multi-Agent Systems for Personalized Gardening

DGX agent

arXiv:2605.23193v1 Announce Type: cross Abstract: Gardening is critical to support well-being, cultural continuity, and food autonomy, yet existing digital tools often provide generic advice that over

local-aiarxiv-cs-cl
25 May 2026
Model Releases

Cultural Adaptation in Large Language Models for Political Discourse

DGX agent

arXiv:2605.23332v1 Announce Type: new Abstract: The integration of large language models into political discourse analysis creates new opportunities for comparative research, policy analysis, and civi

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Decomposing Queries into Tool Calls for Long-Video Keyframe Retrieval

DGX agent

arXiv:2605.23826v1 Announce Type: cross Abstract: Keyframe selection is a direct way to provide verifiable visual evidence for long-video question answering (QA). Queries differ in what they require,

model-releasesarxiv-cs-cl
25 May 2026
Research

DELICATE: Diachronic Entity LInking using Classes And Temporal Evidence

DGX agent

arXiv:2511.10404v2 Announce Type: replace Abstract: In spite of the remarkable advancements in the field of Natural Language Processing, the task of Entity Linking (EL) remains challenging in the fiel

researcharxiv-cs-cl
25 May 2026
Model Releases

DFKI-MLT at SemEval-2026 TASK 7: Steering Multilingual Models Towards Cultural Knowledge

DGX agent

arXiv:2605.23069v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used across diverse linguistic and cultural contexts, yet their cultural knowledge remains uneven across r

model-releasesarxiv-cs-cl
25 May 2026
Safety

Differences in Typological Alignment in Language Models' Treatment of Differential Argument Marking

DGX agent

arXiv:2602.17653v2 Announce Type: replace Abstract: Recent work has shown that language models (LMs) trained on synthetic corpora can exhibit typological preferences that resemble cross-linguistic reg

safetyarxiv-cs-cl
25 May 2026
Applications

Emotion Recognition in Sign Language Conversation

DGX agent

arXiv:2605.23328v1 Announce Type: new Abstract: Emotion Recognition in Conversation is a core component of affective computing, while current resources of sign language emotion datasets primarily focu

applicationsarxiv-cs-cl
25 May 2026
Safety

Entropy-Aware On-Policy Distillation of Language Models

DGX agent

arXiv:2603.07079v2 Announce Type: replace-cross Abstract: On-policy distillation is a promising approach for transferring knowledge between language models, where a student learns from dense token-lev

safetyarxiv-cs-cl
25 May 2026
Safety

EquiSumm : A Gender Bias-Aware Framework for Inclusive Tweet Summarization

DGX agent

arXiv:2605.23412v1 Announce Type: new Abstract: While social media platforms, such as Twitter, provide a medium for large-scale opinion sharing during news events, it is manually impossible for indivi

safetyarxiv-cs-cl
25 May 2026
Research

Evaluating Counterfactual Strategic Reasoning in Large Language Models

DGX agent

arXiv:2603.19167v2 Announce Type: replace Abstract: We evaluate Large Language Models (LLMs) in repeated game-theoretic settings to assess whether strategic performance reflects genuine reasoning or r

researcharxiv-cs-cl
25 May 2026
Applications

Evaluating Customized vs. Generalist Transformer-based Models for Legal Contract Classification

DGX agent

arXiv:2508.07849v2 Announce Type: replace Abstract: Despite advances in legal NLP, no comprehensive evaluation of Transformer-based models customized for legal tasks (referred to as `legal-specific' m

applicationsarxiv-cs-cl
25 May 2026
Model Releases

Evaluating Memory Structure in LLM Agents

DGX agent

arXiv:2602.11243v2 Announce Type: replace-cross Abstract: Modern LLM-based agents and chat assistants rely on long-term memory frameworks to store reusable knowledge, recall user preferences, and augm

model-releasesarxiv-cs-cl
25 May 2026
Safety

Fast-dDrive: Efficient Block-Diffusion VLM for Autonomous Driving

DGX agent

arXiv:2605.23163v1 Announce Type: new Abstract: End-to-end autonomous driving via Vision-Language-Action (VLA) models demands a precarious balance between high-fidelity trajectory planning and efficie

safetyarxiv-cs-cl
25 May 2026
Safety

From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning

DGX agent

arXiv:2605.23382v1 Announce Type: new Abstract: Agentic reinforcement learning (Agentic RL) has achieved strong progress in tasks with clear success signals. However, many real-world agent application

safetyarxiv-cs-cl
25 May 2026
Research

GEMQ: Global Expert-Level Mixed-Precision Quantization for MoE LLMs

DGX agent

arXiv:2605.23078v1 Announce Type: cross Abstract: Mixture-of-Experts Large Language Models (MoE-LLMs) achieve strong performance but incur substantial memory overhead due to massive expert parameters.

researcharxiv-cs-cl
25 May 2026
Local Ai

HawkesLLM: Semantic Uncertainty Propagation in Agentic Text Simulation

DGX agent

arXiv:2605.23043v1 Announce Type: new Abstract: Agentic text-simulation systems write in sequence, with each item becoming possible context for later steps. That makes uncertainty path-dependent: an e

local-aiarxiv-cs-cl
25 May 2026
Research

Hidden Human-Like Nature of Machine-Generated Texts: Theory and Detection Enhancement

DGX agent

arXiv:2605.23190v1 Announce Type: new Abstract: Machine-generated texts (MGTs) produced by large language models (LLMs) are increasingly prevalent across various applications, while their potential mi

researcharxiv-cs-cl
25 May 2026
Model Releases

Hierarchical Concept Geometry in Language Models Emerges from Word Co-occurrence

DGX agent

arXiv:2605.23821v1 Announce Type: new Abstract: We propose a distributional theory of how hypernymy -- the ``is-a'' relation between general and specific concepts -- is encoded geometrically in langua

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

How Far Are We from Generating Missing Modalities with Foundation Models?

DGX agent

arXiv:2506.03530v3 Announce Type: replace-cross Abstract: Multimodal foundation models have demonstrated impressive capabilities across diverse tasks. However, their potential as plug-and-play solutio

model-releasesarxiv-cs-cl
25 May 2026
Applications

How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework

DGX agent

arXiv:2605.23651v1 Announce Type: new Abstract: While factual correctness and task-performance have been in focus of Large Language Model (LLM) research for a long time, the fundamental question of ho

applicationsarxiv-cs-cl
25 May 2026
Research

Improving Sampling for Masked Diffusion Models via Information Gain

DGX agent

arXiv:2602.18176v3 Announce Type: replace Abstract: Masked Diffusion Models (MDMs) enable flexible decoding orders, yet existing samplers remain largely greedy, selecting locally certain tokens withou

researcharxiv-cs-cl
25 May 2026
Research

InfiGFusion: Graph-on-Logits Distillation via Efficient Gromov-Wasserstein for Model Fusion

DGX agent

arXiv:2505.13893v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have intensified efforts to fuse heterogeneous open-source models into a unified system that inherit

researcharxiv-cs-cl
25 May 2026
Applications

Is a Document Educational or Just Wikipedia-Style? -- Pitfalls of Classifier-Based Quality Filtering

DGX agent

arXiv:2605.23721v1 Announce Type: new Abstract: Classifier-based Quality Filtering has recently emerged as a fundamental technique in constructing pre-training corpora. The ability to deploy a single

applicationsarxiv-cs-cl
25 May 2026
Applications

Knowledge Distillation for Low-Resource Open-source Text-to-SQL Model

DGX agent

arXiv:2605.22843v1 Announce Type: new Abstract: Text-to-SQL converts natural language questions into executable SQL queries, enabling non-technical users to access relational databases for analytics a

applicationsarxiv-cs-cl
25 May 2026
Tutorials

Learnability-Informed Fine-Tuning of Diffusion Language Models

DGX agent

arXiv:2605.22939v1 Announce Type: new Abstract: We aim to improve the reasoning capabilities of diffusion language models (DLMs). While SFT is a popular post-training recipe for autoregressive models,

tutorialsarxiv-cs-cl
25 May 2026
Model Releases

Metadata Predictability Is Not Evidence Dependence: An Intervention-Based Audit for Weak-Label Benchmarks

DGX agent

arXiv:2605.23701v1 Announce Type: new Abstract: We study a protocol-level test for weak-label benchmarks: whether benchmark outputs change when the provided evidence is intervened on. Metadata-only sh

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

ModeSwitch-LLM: A Lightweight Phase-Aware Controller for Cross-Mode LLM Inference on a Single GPU

DGX agent

arXiv:2605.23057v1 Announce Type: cross Abstract: ModeSwitch-LLM is a lightweight request-boundary controller for improving single-GPU large language model inference efficiency by routing each request

model-releasesarxiv-cs-cl
25 May 2026
Model Releases

Multi-SpatialMLLM: Multi-Frame Spatial Understanding with Multi-Modal Large Language Models

DGX agent

arXiv:2505.17015v2 Announce Type: replace-cross Abstract: Multi-modal large language models (MLLMs) have rapidly advanced in visual tasks, yet their spatial understanding remains limited to single ima

model-releasesarxiv-cs-cl
25 May 2026
Research

Multilingual Knowledge Transfer under Data Constraints via Lexical Interventions

DGX agent

arXiv:2605.23885v1 Announce Type: new Abstract: Cross-lingual knowledge transfer is critical for building high-performing multilingual language models for languages with insufficient training data. Wh

researcharxiv-cs-cl
25 May 2026
← Previous
1…7980818283…162
Next →