AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-cl”

GridTimelineEvolution
7,700 results
Research

Neuron-based Personality Trait Induction in Large Language Models

DGX agent

arXiv:2410.12327v2 Announce Type: replace Abstract: Large language models (LLMs) have become increasingly proficient at simulating various personality traits, an important capability for supporting re

researcharxiv-cs-cl
11 Jun 2026
Agents
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Notes2Skills: From Lab Notebooks to Certainty-Aware Scientific Agent Skills

DGX agent

arXiv:2606.11897v1 Announce Type: new Abstract: Scientific discovery workflows usually contain and rely heavily on lab notes, where researchers record observations, interpret uncertain results, and pl

agentsarxiv-cs-cl
11 Jun 2026
Research

On The Effectiveness-Fluency Trade-Off In LLM Conditioning: A Systematic Study

DGX agent

arXiv:2606.12234v1 Announce Type: new Abstract: Controlling the output of Large Language Models (LLMs) is a central challenge for their reliable deployment, yet a clear understanding of the involved t

researcharxiv-cs-cl
11 Jun 2026
Safety

One Jailbreak, Many Tongues: Learning Language-Insensitive Intention Representations for Multilingual Jailbreak Detection

DGX agent

arXiv:2606.11202v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in applications for global multilingual users, yet safety training remains concentrated in domina

safetyarxiv-cs-cl
11 Jun 2026
Research

ProHiFlo: Hierarchical Flow Matching with Functional Guidance for De Novo Protein Generation

DGX agent

arXiv:2606.11243v1 Announce Type: cross Abstract: De novo protein generation has transformative potential in therapeutic design, enzyme engineering, and synthetic biology. While diffusion-based and fl

researcharxiv-cs-cl
11 Jun 2026
Model Releases

Reassessing High-Performing LLMs on Polish Medical Exams: True Competence or Bias-Driven Performance?

DGX agent

arXiv:2606.12250v1 Announce Type: new Abstract: Large language models (LLMs) in medicine are mainly evaluated using multiple-choice question answering (MCQA), which can overestimate real clinical abil

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

RLCSD: Reinforcement Learning with Contrastive On-Policy Self-Distillation

DGX agent

arXiv:2606.11709v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) provides dense, token-level supervision for reasoning models by aligning a model's own distribution with the distri

safetyarxiv-cs-cl
11 Jun 2026
Safety

SAGE: Answer-Conditioned Uncertainty Targets for Verbal Uncertainty Alignment

DGX agent

arXiv:2606.11512v1 Announce Type: new Abstract: Large language models increasingly express uncertainty through natural-language statements, yet these expressions often fail to reflect the model's samp

safetyarxiv-cs-cl
11 Jun 2026
Safety

Scenario-based Probing and Steering Cultural Values in Large Language Models--Extended Version

DGX agent

arXiv:2606.11399v1 Announce Type: new Abstract: Large Language Models (LLMs) are deployed across cultural contexts but often reflect homogenized values inherited from training data. Evaluations of cul

safetyarxiv-cs-cl
11 Jun 2026
Safety

Schutzen: Evaluating LLM Safety in Bulgarian and German Contexts

DGX agent

arXiv:2606.11316v1 Announce Type: new Abstract: Large language models are increasingly deployed across professional domains, bringing hard-to-predict risks, including the generation of harmful or disr

safetyarxiv-cs-cl
11 Jun 2026
Research

Semantic Grading of Written Answers in Low-Resource Language Bangla Using a Fine-Tuned Lightweight Language Model

DGX agent

arXiv:2606.11931v1 Announce Type: new Abstract: Bangla is among the world's most widely spoken languages, yet it remains underserved in educational NLP research. In many remote and rural regions, acce

researcharxiv-cs-cl
11 Jun 2026
Model Releases

SoftMatcha 2: A Fast and Soft Pattern Matcher for Trillion-Scale Corpora

DGX agent

arXiv:2602.10908v2 Announce Type: replace Abstract: We present SoftMatcha 2, an ultra-fast and flexible search algorithm that enables search over trillion-scale natural language corpora in under 0.3 s

model-releasesarxiv-cs-cl
11 Jun 2026
Tutorials

SOMA-SQL: Resolving Multi-Source Ambiguity in NL-to-SQL via Synthetic Log and Execution Probing

DGX agent

arXiv:2606.11424v1 Announce Type: new Abstract: Natural language interfaces to databases aim to translate user questions into executable SQL, yet remain brittle in real-world settings where questions

tutorialsarxiv-cs-cl
11 Jun 2026
Research

StanceNakba Shared Task: Actor and Topic-Aware Stance Detection in Public Discourse

DGX agent

arXiv:2606.12068v1 Announce Type: new Abstract: We present StanceNakba 2026, a shared task on stance detection in polarized social media discourse related to the Palestinian-Israeli conflict, organize

researcharxiv-cs-cl
11 Jun 2026
Model Releases

Steering the Noise: Turning Random Perturbations into Effective Descent for Memory-Efficient LLM Fine-Tuning

DGX agent

arXiv:2601.04710v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) achieves strong performance but is often limited by the memory overhead of backpropagation. Zeroth-order (Z

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

Teaching Diffusion to Speculate Left-to-Right

DGX agent

arXiv:2606.11552v1 Announce Type: new Abstract: Large language models (LLMs) achieve remarkable performance across a wide range of tasks, but their autoregressive decoding process incurs substantial i

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

The Language You Ask In: Language-Conditioned Ideological Divergence in LLM Analysis of Contested Political Documents

DGX agent

arXiv:2601.12164v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as analytical tools across multilingual contexts, yet their outputs may carry systemati

model-releasesarxiv-cs-cl
11 Jun 2026
Research

The Long Tail, Not the Front Page: Cold-Start Prediction of Crowd Highlight Salience

DGX agent

arXiv:2606.11654v1 Announce Type: cross Abstract: A social highlighter's most useful signal -- which passages a crowd of readers marks -- exists only for documents people have already read. Can the ag

researcharxiv-cs-cl
11 Jun 2026
Agents

The Periodic Table of LLM Reasoning: A Structured Survey of Reasoning Paradigms, Methods, and Failure Modes

DGX agent

arXiv:2606.11470v1 Announce Type: new Abstract: Large Language Models (LLMs) have achieved strong performance across natural language processing tasks, yet reliable reasoning remains an open challenge

agentsarxiv-cs-cl
11 Jun 2026
Safety

UniReason-Med: A Shared Grounded Reasoning Interface for 2D-to-3D Transfer in Medical VQA

DGX agent

arXiv:2606.11740v1 Announce Type: cross Abstract: We study whether grounded reasoning supervision from abundant 2D medical images can improve 3D medical VQA when both input types are aligned through a

safetyarxiv-cs-cl
11 Jun 2026
Safety

UR-BERT: Scaling Text Encoders for Massively Multilingual TTS Through Universal Romanization and Speech Token Prediction

DGX agent

arXiv:2606.11681v1 Announce Type: new Abstract: We propose UR-BERT, a Romanized transcription-based text-to-speech (TTS) encoder for massively multilingual TTS systems. Conventional grapheme-to-phonem

safetyarxiv-cs-cl
11 Jun 2026
Applications

uva-irlab-conv at SemEval-2026 Task 8: Multi-Turn RAG with Learned Sparse Retrieval and Listwise Reranking

DGX agent

arXiv:2606.11945v1 Announce Type: new Abstract: This report describes our participation in SemEval-2026 Task 8 on multi-turn retrieval and question answering. The task evaluates conversational systems

applicationsarxiv-cs-cl
11 Jun 2026
Research

Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery

DGX agent

arXiv:2602.02726v2 Announce Type: replace-cross Abstract: Large language models (LLMs) encode rich semantic information in their hidden states, yet it remains difficult to understand what information

researcharxiv-cs-cl
11 Jun 2026
Model Releases

Verifiable Environments Are LEGO Bricks: Recursive Composition for Reasoning Generalization

DGX agent

arXiv:2606.12373v1 Announce Type: new Abstract: Reinforcement Learning (RL) with verifiable environments has emerged as a powerful approach for enhancing the reasoning capabilities of Large Language M

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

VietMed-MCQ: A Consistency-Filtered Data Synthesis Framework for Vietnamese Traditional Medicine Evaluation

DGX agent

arXiv:2601.03792v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable proficiency in general medical domains. However, their performance significantly degrades

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

When Does Language Matter? Multilingual Instructions Reveal Step-wise Language Sensitivity in Vision-Language-Action Models

DGX agent

arXiv:2606.11906v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong performance in language-conditioned robotic manipulation, yet their robustness to linguistic varia

model-releasesarxiv-cs-cl
11 Jun 2026
Tutorials

When is Your LLM Steerable?

DGX agent

arXiv:2606.11599v1 Announce Type: new Abstract: Activation steering offers a lightweight approach to control language models' behavior at inference time, but whether it succeeds or fails heavily depen

tutorialsarxiv-cs-cl
11 Jun 2026
Agents

When More Documents Hurt RAG: Mitigating Vector Search Dilution with Domain-Scoped, Model-Agnostic Retrieval

DGX agent

arXiv:2606.11350v1 Announce Type: new Abstract: Retrieval-augmented generation degrades when scaled to large, heterogeneous document collections, where dense similarity loses discriminative power, and

agentsarxiv-cs-cl
11 Jun 2026
Research

Where Do Backdoors Live? A Component-Level Analysis of Backdoor Propagation in Speech Language Models

DGX agent

arXiv:2510.01157v4 Announce Type: replace Abstract: Speech language models (SLMs) are systems of systems: independent components that unite to achieve a common goal. Despite their heterogeneous nature

researcharxiv-cs-cl
11 Jun 2026
Model Releases

Which Models Are Our Models Built On? Auditing Invisible Dependencies in Modern LLMs

DGX agent

arXiv:2606.12385v1 Announce Type: new Abstract: Modern LLM training pipelines increasingly rely on other models to generate data, filter corpora, judge outputs, and guide development decisions. These

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

Which Speech Representation Better Matches Text-Native Reasoning? A Study of Speech-Text Alignment on Frame Rate and Representation

DGX agent

arXiv:2606.12199v1 Announce Type: cross Abstract: Spoken dialogue models typically start from text LLM backbones, yet reasoning often degrades when conditioning on speech instead of text. We attribute

safetyarxiv-cs-cl
11 Jun 2026
Research

A Continuous-Time Markov Chain Framework for Insertion Language Models

DGX agent

arXiv:2606.10199v1 Announce Type: cross Abstract: Insertion Language Models (ILMs) offer several advantages over left-to-right generation and mask-based generation. However, existing formulations of i

researcharxiv-cs-cl
10 Jun 2026
Local Ai

A Navigable Manifold of Hypothesized Consciousness-Spectrum States in Language Model Representations

DGX agent

arXiv:2606.09894v1 Announce Type: cross Abstract: Across contemplative, philosophical, and psychological accounts, human consciousness is often described along a similar spectrum, ranging from reactiv

local-aiarxiv-cs-cl
10 Jun 2026
Research

AI Application Gives Users Real-Time Feedback on the Level of Peace in the Social Media Videos They Watch

DGX agent

arXiv:2601.05232v3 Announce Type: replace Abstract: Most people now get their news from videos on social media, such as YouTube and Facebook, rather than through curated journalism. 'We become what we

researcharxiv-cs-cl
10 Jun 2026
Model Releases

An Industrial-Scale Insurance LLM Achieving Verifiable Domain Mastery and Hallucination Control without Competence Trade-offs

DGX agent

arXiv:2603.14463v2 Announce Type: replace Abstract: Adapting Large Language Models (LLMs) to high-stakes vertical domains like insurance presents a significant challenge: scenarios demand strict adher

model-releasesarxiv-cs-cl
10 Jun 2026
Research

AnimeScore: A Preference-Based Dataset and Framework for Evaluating Anime-Like Speech Style

DGX agent

arXiv:2603.11482v2 Announce Type: replace-cross Abstract: Evaluating 'anime-like' voices currently relies on costly subjective judgments, yet no standardized objective metric exists. A key challenge i

researcharxiv-cs-cl
10 Jun 2026
Research

ArabiGEE: A Hierarchical Taxonomy for Arabic Grammatical Error Explanation

DGX agent

arXiv:2606.10765v1 Announce Type: new Abstract: We introduce ArabiGEE, the first comprehensive Arabic grammatical error explanation (GEE) taxonomy grounded in explicit error types. Unlike existing GEE

researcharxiv-cs-cl
10 Jun 2026
Research

Are We Evaluating Knowledge or Phrasing? Mitigating MCQA Sensitivity with ParaEval

DGX agent

arXiv:2606.10657v1 Announce Type: new Abstract: Multiple-choice (MCQA) benchmarks are the standard for evaluating pretrained large language models, but their reliance on log-likelihood scoring makes t

researcharxiv-cs-cl
10 Jun 2026
Model Releases

Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It

DGX agent

arXiv:2606.11052v1 Announce Type: new Abstract: Chain-of-thought (CoT) supervised fine-tuning (SFT) is widely adopted to improve reasoning ability, yet we find that it systematically degrades long-con

model-releasesarxiv-cs-cl
10 Jun 2026
Safety

Automated Alignment between Elicitation Interviews and Requirements

DGX agent

arXiv:2510.08622v2 Announce Type: replace Abstract: Software requirements are derived from a variety of elicitation techniques, many of which have a conversational nature, like interviews. However, ev

safetyarxiv-cs-cl
10 Jun 2026
Safety

Automated Scoring of Arabic Text Using Large Language Models: A Literature Review

DGX agent

arXiv:2606.09830v1 Announce Type: new Abstract: In modern educational systems, Automatic Text Scoring (ATS) plays a central role by enabling scalable and consistent evaluation of learner responses wit

safetyarxiv-cs-cl
10 Jun 2026
Model Releases

Benchmarking and Exploring the Capabilities of LLMs for Attack Investigations

DGX agent

arXiv:2606.10281v1 Announce Type: cross Abstract: This paper presents AuditBench, a new benchmark dataset for evaluating the capabilities of LLMs at investigating security-related system audit logs. W

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

BenSyc: Benchmarking Conversational Sycophancy and Human Alignment in LLMs for Bengali Contexts

DGX agent

arXiv:2606.10061v1 Announce Type: new Abstract: Large language models (LLMs) increasingly participate in emotionally sensitive social conversations, where responses may shift from balanced support tow

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

Beyond Memorization: Distinguishing Between Pattern-Based and Epistemic Reasoning in LLMs Using Epistemic Puzzles

DGX agent

arXiv:2603.21350v2 Announce Type: replace Abstract: Epistemic reasoning requires agents to infer the state of the world from partial observations and information about other agents' knowledge. Prior w

model-releasesarxiv-cs-cl
10 Jun 2026
Research

CGES: Confidence-Guided Early Stopping for Efficient and Accurate Self-Consistency

DGX agent

arXiv:2511.02603v2 Announce Type: replace Abstract: Large language models (LLMs) are often queried multiple times at test time, with predictions aggregated by majority vote. While effective, this self

researcharxiv-cs-cl
10 Jun 2026
Model Releases

CodeAlchemy: Synthetic Code Rewriting at Scale

DGX agent

arXiv:2606.10087v1 Announce Type: new Abstract: Pre-training on raw code teaches syntax but provides sparse signal for diverse real-world task formats. While synthetic data has proven transformative f

model-releasesarxiv-cs-cl
10 Jun 2026
Model Releases

CommonLID: Re-evaluating State-of-the-Art Language Identification Performance on Web Data

DGX agent

arXiv:2601.18026v2 Announce Type: replace Abstract: Language identification (LID) is a fundamental step in curating multilingual corpora. However, LID models still perform poorly for many languages, e

model-releasesarxiv-cs-cl
10 Jun 2026
Applications

Compiling Rewrite Rules to Finite-State Transducers with the Worsening Trick

DGX agent

arXiv:2606.10059v1 Announce Type: cross Abstract: Finite-state transducers (FSTs) are essential for modeling string rewriting in computational linguistics and natural language processing (NLP), partic

applicationsarxiv-cs-cl
10 Jun 2026
← Previous
1…4748495051…161
Next →