Sens.aiAI signal desk
TodayRadarBriefing
Admin sign in
Sens.aiSensif.aiAI signal desk
ModelsResearchCountriesOpen vs ClosedComputeBetaCommentary
Loading…
TodayRadarBriefingAdmin

Radar

Research Radar

Which research topics are accelerating, and where earlier papers have already landed in shipped models.

Filtered to Other — 75 papersClear filter

More papers match this topic — see “Show more” below.

Papers this week

50

Trending topic

Agents

▲ 19%

papers this window vs prior window

Papers linked to models

128

linked by the desk, past 90 days

Median days paper → model

—

lower is faster

Topic velocity

papers this window by topic · Δ vs prior window · a paper can carry more than one topic

AgentsAgents
134▲19%
Evaluation & BenchmarksEvaluation & Benchmarks
39▲44%
MultimodalMultimodal
23▼30%
Safety & Alignment

Topic momentum

prior window → this window

Agents

134 this window ▲19%

Evaluation & Benchmarks

39 this window ▲44%

Multimodal

23 this window ▼30%

Where research lands: paper → model linkage

Links are AI-inferred by the desk's LLM from tech reports, system cards and citations — each carries a confidence level and is not a claim by the authors.

Let Credit Follow Computation: Architecture-Aware Credit Transport for Large Language Model Reinforcement Learning

arXiv:2608.21501AI-inferred · high

KVBoost: Chunk-Level Key-Value Cache Reuse with Deviation-Guided Recomputation for Efficient Large Language Model Inference

arXiv:2608.21362AI-inferred · high

K-Bench: measuring model performance on real scientific agent requests

arXiv:2608.21601AI-inferred · high

IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents

Filtered to Other — 75 papersClear filter

More papers match this topic — see “Show more” below.

Papers on Other

Clear filter
large language modelsalgorithmic discoverypublished Aug 25, 2026 · arXiv:2608.21584

Data-Driven Dynamic Algorithm Dispatch with Large Language Models

oral cancer screeningedge AIpublished Aug 25, 2026 · arXiv:2608.21583

Robust Lightweight Deep Learning Models for Oral Cancer Screening

Safety & Alignment
22▲47%
Training & OptimizationTraining & Optimization
21▲50%
Efficiency & InferenceEfficiency & Inference
19▼10%
ReasoningReasoning
15▲25%
InterpretabilityInterpretability
12▼20%
Reinforcement LearningReinforcement Learning
10▼29%
Retrieval & RAGRetrieval & RAG
9▼18%
Code GenerationCode Generation
8▼27%
Privacy & SecurityPrivacy & Security
7▲250%
Language UnderstandingLanguage Understanding
5▲25%
Generative ModelsGenerative Models
5▼67%
Robotics & Embodied AIRobotics & Embodied AI
5▼29%
Long ContextLong Context
4▼20%
Speech & AudioSpeech & Audio
4▼33%
Mixture of ExpertsMixture of Experts
3▲50%
Computer VisionComputer Vision
3▼63%
Data & Synthetic DataData & Synthetic Data
1▼75%
OtherOther
165▲12%
View as table
TopicPapers this windowPrior windowΔ
Agents134113+19%
Evaluation & Benchmarks3927+44%
Multimodal2333-30%
Safety & Alignment2215+47%
Training & Optimization2114+50%
Efficiency & Inference1921-9%
Reasoning1512+25%
Interpretability1215-20%
Reinforcement Learning1014-29%
Retrieval & RAG911-18%
Code Generation811-27%
Privacy & Security72+250%
Language Understanding54+25%
Generative Models515-67%
Robotics & Embodied AI57-29%
Long Context45-20%
Speech & Audio46-33%
Mixture of Experts32+50%
Computer Vision38-62%
Data & Synthetic Data14-75%
Other165148+12%

Safety & Alignment

22 this window ▲47%

Training & Optimization

21 this window ▲50%

Efficiency & Inference

19 this window ▼10%
arXiv:2608.06735AI-inferred · high

NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs

arXiv:2608.07167AI-inferred · high

WebGrader: Training LLMs for Web Development with Self-Evolving Programmatic Grader

arXiv:2608.06474AI-inferred · high

Divergent Response Modes in Frontier Language Models Under Steering Pressure

arXiv:2608.06578AI-inferred · high

Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory

arXiv:2608.07169AI-inferred · high

Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills

arXiv:2608.07885AI-inferred · high

SodaMem: Evidence-Grounded Temporal Graph Memory for LLM Agents

arXiv:2608.08055AI-inferred · high

Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation

arXiv:2608.07762AI-inferred · high
Qwen 3.8 Max2 links
claude-fable-5, claude-sonnet-5, claude-opus-52 links
GPT-5.64 links
DeepSeek-V4-Flash-07313 links
GPT-5.6 Luna2 links
GLM-5.31 link
View as table
PaperLinked modelRelationConfidence
Let Credit Follow Computation: Architecture-Aware Credit Transport for Large Language Model Reinforcement Learning (2608.21501)Qwen 3.8 Maxevaluates99%
KVBoost: Chunk-Level Key-Value Cache Reuse with Deviation-Guided Recomputation for Efficient Large Language Model Inference (2608.21362)Qwen 3.8 Maxevaluates98%
K-Bench: measuring model performance on real scientific agent requests (2608.21601)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates100%
K-Bench: measuring model performance on real scientific agent requests (2608.21601)GPT-5.6evaluates100%
IB-RL: Isolated Bilateral Reinforcement Learning for Strategic Dialogue Agents (2608.06735)DeepSeek-V4-Flash-0731evaluates95%
NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs (2608.07167)GPT-5.6evaluates98%
WebGrader: Training LLMs for Web Development with Self-Evolving Programmatic Grader (2608.06474)GPT-5.6evaluates99%
Divergent Response Modes in Frontier Language Models Under Steering Pressure (2608.06578)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates99%
Divergent Response Modes in Frontier Language Models Under Steering Pressure (2608.06578)GPT-5.6evaluates99%
Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory (2608.07169)GPT-5.6 Lunatechnique used98%
Reason Wide, Not Deep: Amortizing the Reasoning Premium into Distilled Skills (2608.07885)GPT-5.6 Lunaevaluates98%
SodaMem: Evidence-Grounded Temporal Graph Memory for LLM Agents (2608.08055)DeepSeek-V4-Flash-0731technique used98%
Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation (2608.07762)DeepSeek-V4-Flash-0731evaluates98%
Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation (2608.07762)GLM-5.3evaluates98%
Who Verifies the Benchmark? Decentralizing Trust in Large Language Model Evaluation (2608.07762)GPT-5.6 Lunaevaluates98%
An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Glaucoma Detection from Fundus Photography (2608.07651)GPT-5.6 Lunaevaluates99%
An Agentic AI Framework Overcomes Fundamental Limitations of Large Language Models for Glaucoma Detection from Fundus Photography (2608.07651)Gemini 3.7 Flashevaluates99%
When the Judge Should Not Decide: Evidence-Locked, Non-Compensatory Selection Bounds LLM-Judge Failure in Reasoning Pipelines (2608.07813)DeepSeek-V4-Flash-0731evaluates95%
The Authority Expectancy Effect in Multi-User Conflict (2608.08026)Grok 4.6evaluates95%
The Authority Expectancy Effect in Multi-User Conflict (2608.08026)GPT-5.6 Lunaevaluates95%
The Authority Expectancy Effect in Multi-User Conflict (2608.08026)Gemini 3.7 Flashevaluates95%
The Authority Expectancy Effect in Multi-User Conflict (2608.08026)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates95%
Can Legal AI Know When It Is Wrong? And Do Students Know When It Is? (2608.21089)GPT-5.6 Lunaevaluates99%
Calibrating Criterion Revision in LLM Agents: Failure Modes and a Trace-Anchored Protocol (2608.20729)Qwen 3.8 Maxevaluates98%
Why2Speak: Faithful Reasoning for Abstaining Action Policies (2608.20670)Qwen 3.8 Maxevaluates98%
No Judgment Without a Reason: Counterfactual Receipts for Versioned AI Evaluators (2608.20938)Qwen 3.8 Maxevaluates99%
Automated Trajectory Evaluation for Mobile Agents via Step-Level Consequence Reasoning and Aggregation (2608.20797)Qwen 3.8 Maxtechnique used98%
Beyond Endpoint Gains: A Weight-Delta Audit of Medical Specialization (2608.20768)GPT-5.6 Lunaevaluates99%
Beyond Endpoint Gains: A Weight-Delta Audit of Medical Specialization (2608.20768)Qwen 3.8 Maxevaluates99%
Beyond Endpoint Gains: A Weight-Delta Audit of Medical Specialization (2608.20768)Gemma 4evaluates99%
Natural-Language-Guided Generator-Agnostic Shortlisting for Protein Binder Design (2608.20755)GPT-5.6 Lunatechnique used95%
DirEAG: Dirichlet Evidence Aggregation for Calibrating Verbalized Confidence in Mathematical Reasoning (2608.20717)Gemma 4evaluates95%
DirEAG: Dirichlet Evidence Aggregation for Calibrating Verbalized Confidence in Mathematical Reasoning (2608.20717)Mistral OCR 4evaluates95%
DirEAG: Dirichlet Evidence Aggregation for Calibrating Verbalized Confidence in Mathematical Reasoning (2608.20717)Qwen 3.8 Maxevaluates95%
Nexus: Depth-Adaptive KV-Cache Splicing and Retrieval-Decoupled Tool Routing for Agentic LLMs on Unified Memory (2608.20397)Qwen 3.8 Maxevaluates99%
Truth Lies Deep: Countering Semantic Camouflage via Latent Intent Verification (2608.20378)Gemma 4evaluates98%
Truth Lies Deep: Countering Semantic Camouflage via Latent Intent Verification (2608.20378)Qwen 3.8 Maxevaluates98%
PrimeAgentOrchestrator: Memory-Primed Agent Spawning for Personal AI Infrastructure (2608.20342)claude-fable-5, claude-sonnet-5, claude-opus-5technique used98%
Personalized Privacy Control in LLMs via Attention Head Intervention (2608.21209)Qwen 3.8 Maxevaluates99%
TreeWY: Speculative Verification for Gated DeltaNet Hybrids (2608.20961)Qwen 3.8 Maxevaluates98%
Weighted Memory Tree: Remembering What Matters for Long-Horizon LLM Agents (2608.20631)Gemma 4evaluates99%
Weighted Memory Tree: Remembering What Matters for Long-Horizon LLM Agents (2608.20631)Qwen 3.8 Maxevaluates99%
FlavourBench: Ranking Frontier Language Models with Executable Culinary Ground Truth (2608.20574)Grok 4.6evaluates99%
StateSight: Benchmarking Latent Spatial-State Reconstruction in Vision-Language Models (2608.20414)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates99%
StateSight: Benchmarking Latent Spatial-State Reconstruction in Vision-Language Models (2608.20414)GPT-5.6 Lunaevaluates99%
SkillLens: Visual Skill Cards for Retrieval-Augmented GUI Action Prediction and On-Policy Distillation (2608.10775)GPT-5.6 Lunaevaluates98%
CHORUS: Complementary Experts for High-Coverage Testbench Stimulus Generation (2608.10090)DeepSeek-V4-Flash-0731evaluates98%
DSAgentBench: Can Agents Automate End-to-End Data-Science Workflows in Real Computer Environments? (2608.10366)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates98%
Apodex Discovery: Reality Benchmarks and Environments for Evaluating and Building Discoverative Artificial Intelligence (2608.11341)GPT-5.6 Lunaevaluates98%
AutoWorldModel-Bench: A State-Centric Benchmark for Automated World-Model Research (2608.11216)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates98%
MBA: Multimodal Benchmark and Agents for Real-World Business Ideation (2608.11616)GPT-5.6 Lunatechnique used98%
When Self-Consistency Backfires: Majority Vote Hurts the Majority of Hard Science Problems for Small LLMs (2608.11403)Qwen 3.8 27Bevaluates99%
XBridge: Entity-Grounded Latent Bridge for Heterogeneous LLM Communication (2608.11676)Qwen 3.8 27Bevaluates95%
Graph-Structured Rubrics: Compiling Rubrics into Typed Evaluation Graphs for LLM Judges (2608.12097)GPT-5.6 Lunaevaluates95%
HUGIN: Enhancing Vision-Language Planning for Autonomous Logistics Sorting (2608.11692)Qwen 3.8 27Bevaluates99%
Claim-Level Reliability Assessment for Efficient Test-Time Reasoning (2608.11994)GPT-5.6 Lunaevaluates95%
From Numbers to Judgment: Specialist LLM Agents and Reinforcement Learning for European Listed Real Estate (2608.11381)Qwen 3.8 27Btechnique used98%
Retrofitting Recurrent Depth into a Pretrained Language Model: Installation, Extrapolation, Transfer, and Retention at Two Parameter Budgets (2608.11233)Qwen 3.8 27Btechnique used99%
Backtrader-Bench: Benchmarking LLM Agents on Algorithmic Trading with Self-Generated MCQs (2608.11232)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates95%
Backtrader-Bench: Benchmarking LLM Agents on Algorithmic Trading with Self-Generated MCQs (2608.11232)GPT-5.6 Lunaevaluates95%
FrontierFinance: A Challenging Benchmark for Measuring Frontier Intelligence of Finance Agents (2608.11683)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates98%
Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence (2608.12036)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates86%
Harnessing agent memory to build lifelong AI partners for materials scientists (2608.11224)GPT-5.6 Lunaevaluates95%
Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Text-to-Image Retrieval (2608.11343)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates99%
Can Frontier LLMs Match Natively Multimodal Embeddings? A Comparison on Hard-Negative Text-to-Image Retrieval (2608.11343)GPT-5.6 Lunaevaluates99%
An Agentic Workflow for Legacy HPC Modernization: Converting the Two-Electron-Integral Core of GAMESS (2608.12249)claude-fable-5, claude-sonnet-5, claude-opus-5technique used95%
Distribird: Literature-Informed Prior Distribution Design for Bayesian Model Calibration (2608.11210)Qwen 3.8 27Bevaluates95%
From Monolithic to Modular: Segment-level Automatic Prompt Optimization (2608.11219)GPT-5.6 Lunaevaluates98%
RecSys Factory: Bounding LLM Agent Autonomy to Decision Points in the Industrial Recommender Lifecycle (2608.11241)claude-fable-5, claude-sonnet-5, claude-opus-5technique used90%
Reasoning Jury: Multi-Model Consensus for Evaluating Reasoning Traces (2608.12585)GPT-5.6 Lunaevaluates95%
Enhancing Virtual Agents through SLMs and Edge-Computing: An Exploratory Evaluation of Think and Memory Processes (2608.13420)Qwen 3.8 27Bevaluates98%
Rethinking Normalization Placement for LLMs: Post-Norm under Curriculum Depth Growing (2608.13156)Qwen 3.8 27Bevaluates95%
Jointly Predicting Courses and Grades Using a Transformer-Based Model (2608.13409)trace-supervised symbolic neural CPUtechnique used99%
Explanatory Engagement Under Rare Anomalous Failure: Asymptotic Rarity in Model Behavior (or: The Asymptotic AI) (2608.13063)Qwen 3.8 27Bevaluates99%
ε-MemEvo: Adaptive Cross-Task Memory Transfer for LLM Program Evolution (2608.12522)AlphaEvolverelated94%
ε-MemEvo: Adaptive Cross-Task Memory Transfer for LLM Program Evolution (2608.12522)GPT-5.6 Lunaevaluates98%
Governed Persistent Memory: Source-Bound State Semantics and Fail-Closed Release for Long-Horizon Agents (2608.12476)Qwen 3.8 27Bevaluates99%
Privacy-Preserving RAG by Concealing Sensitive Information from External LLMs (2608.12675)Qwen 3.8 27Bevaluates95%
Polish Medical Visual Question Answering: Vision-Language Models Underutilize Visual Evidence (2608.12928)GPT-5.6 Lunaevaluates99%
From Passive Delegates to Strategic Negotiators: Reinforcing Social Reasoning in Small Language Models with SocialRL (2608.13787)GPT-5.6 Lunaevaluates95%
Explanation Multiplicity: Circuit-Level Interpretability Evidence Does Not Survive Defensible Analytic Variation (2608.13754)GPT-5.6 Lunaevaluates99%
Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages (2608.14375)GPT-5.6 Lunaevaluates95%
MemoryLake on MemoryArena: A Matched Study of Agent Memory Backends (2608.13883)GPT-5.6 Lunaevaluates95%
Implementing Computational Law in Wolfram Language for the Governance of Artificial Intelligence (2608.13958)GPT-5.6 Lunaevaluates100%
Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning (2608.14290)Qwen 3.8 Maxrelated95%
StateM: Reaching 95.3% Raw Accuracy, or a $15 Frontier Run, on Terminal-Bench 2.1 via Harness Scaling (2608.15089)DeepSeek-V4-Flash-0731evaluates99%
StateM: Reaching 95.3% Raw Accuracy, or a $15 Frontier Run, on Terminal-Bench 2.1 via Harness Scaling (2608.15089)GPT-5.6 Lunaevaluates95%
Large Language Models Show Metacognitive Sensitivity in Medical Reasoning (2608.14552)GPT-5.6 Lunaevaluates99%
When Agentic Executions Fail: Detecting and Localizing Runtime Faults from Telemetry (2608.14680)DeepSeek-V4-Flash-0731evaluates98%
SKILL: Self-correcting Knowledge-guided Iterative Large Language Model Agent for Logic Optimization (2608.14579)GPT-5.6 Lunatechnique used98%
ACTS-SQL: Agentic and Critic-Oriented Tree-Structured SQL Correctness with Large Language Models (2608.15145)GPT-5.6 Lunaevaluates95%
The Unwritten Benchmark: A New Challenge for Multimodal Machine Learning in Abstract Perceptual Reasoning (2608.14558)GPT-5.6 Lunaevaluates98%
The Hallucination Snowball: Modeling Error Propagation as State Transitions in Multi-Agent LLM Pipelines (2608.14588)GPT-5.6 Lunaevaluates98%
LLMs Can Predict Failure Risk, But Struggle to Predict Which Collaboration Protocol Pays Off: Cost-Aware Protocol Routing Across Reasoning Tasks (2608.14927)GPT-5.6 Lunaevaluates95%
Mechanistic Tomography: Designed Measurement for Control-Oriented Interpretability (2608.19338)Qwen 3.8 Maxevaluates98%
Mechanistic Tomography: Designed Measurement for Control-Oriented Interpretability (2608.19338)GPT-5.6 Lunaevaluates98%
When AI Writes, Who Gets Cited? Evidence of Citation Monoculture Across Language Models (2608.19230)GPT-5.6 Lunaevaluates98%
MidTool: Mid-training Data Synthesis for Agentic Tool Use (2608.20314)Qwen 3.8 Maxtechnique used98%
Automatic bioinformatic software named entity recognition from literature (2608.19201)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates90%
Automatic bioinformatic software named entity recognition from literature (2608.19201)Grokevaluates90%
Automatic bioinformatic software named entity recognition from literature (2608.19201)Gemini 3.7 Flashevaluates90%
Automatic bioinformatic software named entity recognition from literature (2608.19201)GPT-5.6 Lunaevaluates90%
Can Conversational AI loosen Us-Versus-Them Boundaries? The Effects of Common, Dual, and Separate Identity Framings on Pro-Immigrant Intergroup Helping (2608.19220)GPT-5.6 Lunatechnique used99%
A Virtual Member of a Community of Practice for the Society of Petroleum Engineers: From Prototype to Deployment (2608.19199)Athena-Brain-8Bevaluates98%
Phantom Gains: Auditing Self-Improvement Against a Measured Null (2608.20290)Qwen 3.8 Maxevaluates98%
ReguSim: Evaluating LLM Agent Rule Grounding in Financial Compliance (2608.19974)Gemini 3.7 Flashevaluates95%
ReguSim: Evaluating LLM Agent Rule Grounding in Financial Compliance (2608.19974)DeepSeek-V4-Flash-0731evaluates95%
PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents (2608.19861)Gemini 3.7 Flashevaluates99%
PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents (2608.19861)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates99%
PolicyGuide: From Guarding One Action to Guiding the Whole Workflow for Policy-Compliant LLM Agents (2608.19861)GPT-5.6 Lunaevaluates99%
SAPO: Single-Rollout Autoregressive Policy Optimization for Agentic Reinforcement Learning (2608.19842)Qwen 3.8 Maxevaluates98%
Can Agent Memory Systems Track Evolving State? (2608.19652)Qwen 3.8 Maxevaluates98%
Can Agent Memory Systems Track Evolving State? (2608.19652)DeepSeek-V4-Flash-0731evaluates98%
From Retrieved Context to Runtime Control: Adaptive Compression for Edge-based RAG (2608.19535)Qwen 3.8 Maxevaluates99%
Air Traffic Control Using Large Language Models: Prompt Engineering, Architecture, and Evaluation (2608.19299)GPT-5.6 Lunaevaluates80%
When Personalization Becomes Bias: Structural and Discursive Religious Framing in AI-Generated Financial Advice (2608.16909)Grokevaluates98%
Different Facets of Verbalised Overconfidence: an Interpretability Study (2608.18106)Qwen 3.8 Maxevaluates99%
DeepTCM1.0: A Multi-Expert AI Agent for Deciphering Mechanisms of Chinese Herbal Formulae Based on General Large Language Models (2608.18103)DeepSeek V4-Flashtechnique used99%
Abliteration Mitigation via Refusal Aliases (2608.18093)Gemma 4evaluates99%
Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities (2608.18090)Gemma 4evaluates85%
Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities (2608.18090)Qwen 3.8 Maxevaluates85%
Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities (2608.18090)Mistral OCR 4evaluates85%
Latent Space Refusal Anchoring for Low-Resource African Languages: Mechanistic Safety Recovery Without Retraining (2608.18089)Qwen 3.8 Maxevaluates99%
Latent Space Refusal Anchoring for Low-Resource African Languages: Mechanistic Safety Recovery Without Retraining (2608.18089)Mistral OCR 4evaluates99%
Beyond the Transcript: Detecting Covert Co ordination in Latent Multi-Agent Communication (2608.19161)Qwen 3.8 Maxevaluates98%
Can a Lightweight Multimodal Model Estimate LLM Reasoning Performance? A Study for Compute-Optimal Document Inference (2608.18591)Qwen 3.8 Maxtechnique used100%
A Jagged Frontier: Evaluating Robustness of Code Agents to Semantics-Preserving Transformations (2608.18389)Qwen 3.8 Maxevaluates98%
A Jagged Frontier: Evaluating Robustness of Code Agents to Semantics-Preserving Transformations (2608.18389)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates98%
Governance Records as Supervision: Verifier-Selected Self-Training for Structured Workflow Repair (2608.18324)Qwen 3.8 Maxevaluates98%
Efficient Adaptation of LLMs for Hate Speech Detection in Low-Resource Languages: A Comparative Study on Roman Urdu (2608.18142)Mistral OCR 4evaluates99%
Position: Collusion Risks Among AI Reasoning Agents Justify Certification Requirements for Making Market Decisions (2608.18078)DeepSeek V4-Flashevaluates95%
Computational Orientalism: Measuring Structural Discourse Bias in Large Language Models Using the Middle East Cultural Sensitivity Score (MECSS) (2608.18100)GPT-5.6 Lunaevaluates99%
Breaking the weakest link to evade vision language models (2608.18938)Qwen 3.8 Maxevaluates99%
Training-Free Inference-Time Self-Reflection and Cost-Bounded Early Stopping for Large Language Models (2608.18884)Qwen 3.8 Maxevaluates98%
CTIFoundry: An Agent-Native Corpus Scaffold for Cyber Threat Intelligence (2608.18613)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates80%
FM-Bench: A Benchmark for Long-Horizon Management with Competing Agents (2608.18423)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates100%
Measuring the Partial-Credit Gap: A Strict Benchmark on Vietnam's 2025 Convex Marking Scheme (2608.18336)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates90%
Measuring the Partial-Credit Gap: A Strict Benchmark on Vietnam's 2025 Convex Marking Scheme (2608.18336)Qwen 3.8 Maxevaluates98%
Solving Is Not Drawing: A Benchmark for Diagrammatic Reasoning in Olympiad Geometry (2608.18111)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates80%
Solving Is Not Drawing: A Benchmark for Diagrammatic Reasoning in Olympiad Geometry (2608.18111)GPT-5.6 Lunaevaluates80%
ComponentBench: Diagnosing Component-Level Failures in Computer-Use Agents (2608.18307)Qwen 3.8 Maxevaluates99%
ComponentBench: Diagnosing Component-Level Failures in Computer-Use Agents (2608.18307)Gemini 3.7 Flashevaluates99%
ComponentBench: Diagnosing Component-Level Failures in Computer-Use Agents (2608.18307)GPT-5.6 Lunaevaluates99%
Potential of ChatGPT in predicting stock market trends based on Twitter Sentiment Analysis (2311.06273)GPT-5.6 Lunaevaluates95%
StagedWorkspace: A Versioned Workspace for Knowledge-Work Agents (2608.18050)GPT-5.6 Lunaevaluates98%
StagedWorkspace: A Versioned Workspace for Knowledge-Work Agents (2608.18050)Gemini 3.7 Flashevaluates98%
Auditing Self-Evolution in Financial Agents: Capability Gains, Security Drift, and Execution-Interface Mismatch (2608.17684)Qwen 3.8 Maxevaluates98%
The Price of Thinking: Reasoning Effort as a Model-Specific API Contract (2608.16956)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates95%
Beyond the Trace: Coupling an Interpretable Reasoning-State Readout to Native MoE Routing (2608.17638)GPT-5.6 Lunaevaluates95%
TRUSS: Towards Task-Reliable and User-Safe Automated Agent Skill Generation (2608.17588)GPT-5.6 Lunaevaluates95%
Agent Lightning v1.0: Towards Harnessed Agentic RL (2608.17528)Qwen 3.8 Maxevaluates99%
LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents (2608.17393)Qwen 3.8 Maxevaluates99%
TileMix: Tile-Centric Mixed-Precision Attention for LLM Inference Acceleration (2608.17336)Qwen 3.8 Maxevaluates98%
SignalReasoner: Assessing the Upper Bound of 3B Models for Signal Mathematical Reasoning (2608.17301)Qwen 3.8 Maxevaluates98%
When Personalization Becomes Bias: Structural and Discursive Religious Framing in AI-Generated Financial Advice (2608.16909)Grok 4.6evaluates99%
When Personalization Becomes Bias: Structural and Discursive Religious Framing in AI-Generated Financial Advice (2608.16909)Gemini 3.7 Flashevaluates98%
When Personalization Becomes Bias: Structural and Discursive Religious Framing in AI-Generated Financial Advice (2608.16909)GPT-5.6 Lunaevaluates98%
FedPref: Federated Preference Learning for Structured Radiology Report Extraction (2608.16971)Qwen 3.8 Maxtechnique used99%
Explicit State Elicitation Is Not Enough: A Controlled Audit of Memory-Policy Classification (2608.17247)GPT-5.6 Lunaevaluates98%
GxP-Agent: Process-DAG Topology for Reliable Clinical Trial Programming with LLM Agents (2608.16890)GPT-5.6 Lunaevaluates99%
GxP-Agent: Process-DAG Topology for Reliable Clinical Trial Programming with LLM Agents (2608.16890)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates99%
StateM: Reaching 95.3% Raw Accuracy, or a $15 Frontier Run, on Terminal-Bench 2.1 via Harness Scaling (2608.15089)DeepSeek V4-Flashevaluates99%
StateM: Reaching 95.3% Raw Accuracy, or a $15 Frontier Run, on Terminal-Bench 2.1 via Harness Scaling (2608.15089)GPT-5.6 Solevaluates99%
Does a Tool Result Carry More Authority Than Plain Text? Three Prospective Studies of False-Claim Adoption in a Synthetic Assignment Task with Claude Opus 5 (2608.14992)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates100%
Small Models Scout Bottleneck Order for Large-Model Data Control (2608.14936)Qwen 3.8 Maxevaluates95%
LLMs Can Predict Failure Risk, But Struggle to Predict Which Collaboration Protocol Pays Off: Cost-Aware Protocol Routing Across Reasoning Tasks (2608.14927)GPT-5.6 Soltechnique used98%
The Hallucination Snowball: Modeling Error Propagation as State Transitions in Multi-Agent LLM Pipelines (2608.14588)Qwen 3.8 Maxevaluates98%
The Hallucination Snowball: Modeling Error Propagation as State Transitions in Multi-Agent LLM Pipelines (2608.14588)GPT-5.6 Solevaluates99%
SKILL: Self-correcting Knowledge-guided Iterative Large Language Model Agent for Logic Optimization (2608.14579)Gemini 3.7 Flashtechnique used98%
SKILL: Self-correcting Knowledge-guided Iterative Large Language Model Agent for Logic Optimization (2608.14579)claude-fable-5, claude-sonnet-5, claude-opus-5technique used98%
SKILL: Self-correcting Knowledge-guided Iterative Large Language Model Agent for Logic Optimization (2608.14579)GPT-5.6 Soltechnique used99%
Anatomy of a Quantized Agent: VRAM Stability and Forecasting in Code-Synthesis Agentic Workloads (2608.15117)Qwen 3.8 Maxevaluates99%
When Agentic Executions Fail: Detecting and Localizing Runtime Faults from Telemetry (2608.14680)DeepSeek V4-Flashevaluates99%
ACTS-SQL: Agentic and Critic-Oriented Tree-Structured SQL Correctness with Large Language Models (2608.15145)GPT-5.6 Solevaluates95%
ACTS-SQL: Agentic and Critic-Oriented Tree-Structured SQL Correctness with Large Language Models (2608.15145)GPT-5.6 Soltechnique used95%
The Unwritten Benchmark: A New Challenge for Multimodal Machine Learning in Abstract Perceptual Reasoning (2608.14558)Gemini 3.7 Flashevaluates98%
The Unwritten Benchmark: A New Challenge for Multimodal Machine Learning in Abstract Perceptual Reasoning (2608.14558)GPT-5.6 Solevaluates98%
Large Language Models Show Metacognitive Sensitivity in Medical Reasoning (2608.14552)GPT-5.6 Solevaluates99%
MemoryLake on MemoryArena: A Matched Study of Agent Memory Backends (2608.13883)GPT-5.6 Soltechnique used95%
From Passive Delegates to Strategic Negotiators: Reinforcing Social Reasoning in Small Language Models with SocialRL (2608.13787)GPT-5.6 Solevaluates99%
Intern-S2-Mobius: Foundation Model with Decoupled Knowledge and Reasoning (2608.14290)Qwen 3.8 Maxtechnique used85%
Implementing Computational Law in Wolfram Language for the Governance of Artificial Intelligence (2608.13958)GPT-5.6 Solevaluates99%
Explanation Multiplicity: Circuit-Level Interpretability Evidence Does Not Survive Defensible Analytic Variation (2608.13754)GPT-5.6 Solevaluates98%
Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages (2608.14375)Gemma 4evaluates95%
Wrong but Useful: Trajectory Value Beyond Answer Correctness in Multi-Agent Messages (2608.14375)GPT-5.6 Solevaluates98%
A Calibrated Test of Internal Action Maps: State Signals Without Global Affine Closure (2608.13626)Qwen 3.8 Maxevaluates98%
Depth-Aware Sensitivity Analysis of Mixture-of-Experts Models via Magnitude-Based Expert Masking (2608.13565)Qwen 3.8 Maxevaluates98%
OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality (2608.05263)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates90%
NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs (2608.07167)GPT-5.6 Solevaluates99%
ε-MemEvo: Adaptive Cross-Task Memory Transfer for LLM Program Evolution (2608.12522)AlphaEvolvecites88%
ε-MemEvo: Adaptive Cross-Task Memory Transfer for LLM Program Evolution (2608.12522)GPT-5.6 Solevaluates98%
Don't Want Your LLM to Recommend Nuclear Strike? Try Asking It in Japanese (2608.12373)Gemini 3.7 Flashevaluates98%
Don't Want Your LLM to Recommend Nuclear Strike? Try Asking It in Japanese (2608.12373)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates98%
Privacy-Preserving RAG by Concealing Sensitive Information from External LLMs (2608.12675)Qwen 3.8 Maxevaluates98%
Rethinking Normalization Placement for LLMs: Post-Norm under Curriculum Depth Growing (2608.13156)Qwen 3.8 Maxevaluates98%
Explanatory Engagement Under Rare Anomalous Failure: Asymptotic Rarity in Model Behavior (or: The Asymptotic AI) (2608.13063)Mistral OCR 4evaluates99%
Explanatory Engagement Under Rare Anomalous Failure: Asymptotic Rarity in Model Behavior (or: The Asymptotic AI) (2608.13063)Qwen 3.8 Maxevaluates99%
Enhancing Virtual Agents through SLMs and Edge-Computing: An Exploratory Evaluation of Think and Memory Processes (2608.13420)Qwen 3.8 Maxevaluates98%
Reasoning Jury: Multi-Model Consensus for Evaluating Reasoning Traces (2608.12585)Gemini 3.7 Flashevaluates95%
Reasoning Jury: Multi-Model Consensus for Evaluating Reasoning Traces (2608.12585)claude-fable-5, claude-sonnet-5, claude-opus-5evaluates95%
knowledge graphsSHACL validationpublished Aug 25, 2026 · arXiv:2608.21418

Composable Trust Infrastructure for Manufacturing Knowledge Graphs: Cross-System Provenance, Temporal Reasoning, and Decision Traceability

root cause analysisdatacenter networkspublished Aug 25, 2026 · arXiv:2608.21412

The Abstention Protocol: RCA for Clos Fabrics

student burnoutstudy habit trackingpublished Aug 25, 2026 · arXiv:2608.21379

RIACT: A Responsible AI System for Personalized Study Habit Tracking and Early Burnout Signal Detection in University Students

AI governanceruntime decision auditingpublished Aug 25, 2026 · arXiv:2608.21363

AIREP: A Protocol for Per-Decision Evidence in AI Runtime Governance

domain adaptationgeographic domain shiftpublished Aug 25, 2026 · arXiv:2608.21567

Quantifying geographic domain shift to decouple the geospatial transferability of human mobility flow generation models

CLIPstreet-view imagerypublished Aug 25, 2026 · arXiv:2608.21761

What Does CLIP Learn for Regional Geolocalization? Probing Visual Cues and Scene Configuration After Adaptation

large language modelslegal AIpublished Aug 24, 2026 · arXiv:2608.21089

Can Legal AI Know When It Is Wrong? And Do Students Know When It Is?

LinkedGPT-5.6 Luna
flow matchingtransition-state predictionpublished Aug 24, 2026 · arXiv:2608.20869

ReCurveflow: A Flow Matching Framework that Learns Curved Reaction Trajectories to Predict Transition State Geometries

causal foundation modelspartial causal identificationpublished Aug 24, 2026 · arXiv:2608.20841

Foundation Models for Partial Causal Identification

graph neural networkscontinuous-time quantum walkspublished Aug 24, 2026 · arXiv:2608.20738

Continuous-Time Quantum Walks based Graph Neural Network

continual learningcatastrophic forgettingpublished Aug 24, 2026 · arXiv:2608.21044

Socialized Division and Collaboration: Rethinking Class-Incremental Learning under Optimization Conflicts

soft tissue simulationfinite element methodpublished Aug 24, 2026 · arXiv:2608.20967

Generalizing Soft Tissue Deformation and Force Prediction Across Material Stiffness and Geometry

physics-informed learningshape-constrained learningpublished Aug 24, 2026 · arXiv:2608.21059

The Cost of a Physics Prior Is Bounded by the Ablation Gap

SQLdatabase systemspublished Aug 24, 2026 · arXiv:2608.20630

SAGE: A Unified Algebra and Self-Adaptive Execution for AI Functions in SQL

semantic IDsautoregressive recommendationpublished Aug 24, 2026 · arXiv:2608.20611

Difficulty-Aware Semantic-ID Optimization for Generative Recommendation

model evaluationprediction certificationpublished Aug 24, 2026 · arXiv:2608.20825

Prediction certification cannot replace explanation certification: a competence envelope for trustworthy AI under compound stress

motion forecastingBayesian uncertainty estimationpublished Aug 24, 2026 · arXiv:2608.20802

SPARC: Single-Pass Scaling for Motion Forecasting with Conformal Bayesian Last Layers

multimodal learningvideo understandingpublished Aug 24, 2026 · arXiv:2608.20958

TLive-Omni: An Omni-Modal Understanding Model for E-Commerce Live Streaming

machine unlearningclaim-level knowledge removalpublished Aug 24, 2026 · arXiv:2608.20960

Can Scientific Claims Be Removed from Large Language Models? A Systematic Evaluation of Claim-Level Unlearning

model specializationweight-delta analysispublished Aug 24, 2026 · arXiv:2608.20768

Beyond Endpoint Gains: A Weight-Delta Audit of Medical Specialization

LinkedGPT-5.6 LunaQwen 3.8 MaxGemma 4
large language modelsprotein designpublished Aug 24, 2026 · arXiv:2608.20755

Natural-Language-Guided Generator-Agnostic Shortlisting for Protein Binder Design

LinkedGPT-5.6 Luna
AI ethicsAI governancepublished Aug 24, 2026 · arXiv:2608.20490

Lost in Translation: How Universal Ethical Values Fail to Translate Across Global Contexts

agentic AIAI adoptionpublished Aug 24, 2026 · arXiv:2608.20425

Who Delegates to AI? Evidence from 53,000 Agent Configurations

artificial phenomenologycategorical mathematicspublished Aug 24, 2026 · arXiv:2608.20420

Categorical AI phenomenology: A first-person approach

IT change managementrisk assessmentpublished Aug 24, 2026 · arXiv:2608.21203

SENTRY: Deterministic, Intelligent Risk Assessment for IT Change Management

attention mechanismstoken presencepublished Aug 24, 2026 · arXiv:2608.21174

From Attention Masks to Inert Zero-Vector Tokens: OAttention and O-Closure for Token Dynamics

difference graph discoverycausal inferencepublished Aug 24, 2026 · arXiv:2608.21117

Root cause analysis via difference graph discovery from linear time-series data

multi-agent LLM systemsKV-cache translationpublished Aug 24, 2026 · arXiv:2608.20617

Dual-Cache Latent Space Communication between Heterogeneous Language Models

neural operatorsphysics-informed machine learningpublished Aug 24, 2026 · arXiv:2608.20477

STCO: Conditional Neural Operators for Time-Dependent PDEs

spiking neural networks3D point cloud recognitionpublished Aug 21, 2026 · arXiv:2608.19232

Active Spiking Perception: The Membrane Potential as a Belief State for Anytime 3D Point Cloud Recognition

language modelsliterature-search agentspublished Aug 21, 2026 · arXiv:2608.19230

When AI Writes, Who Gets Cited? Evidence of Citation Monoculture Across Language Models

LinkedGPT-5.6 Luna
AI regulationsystemic risk assessmentpublished Aug 21, 2026 · arXiv:2608.19278

Mapping General-Purpose AI Governance in Twenty AI Middle-Power Jurisdictions

bioinformaticsnamed entity recognitionpublished Aug 21, 2026 · arXiv:2608.19201

Automatic bioinformatic software named entity recognition from literature

Linkedclaude-fable-5, claude-sonnet-5, claude-opus-5GrokGemini 3.7 FlashGPT-5.6 Luna
uncertainty quantificationconfidence estimationpublished Aug 21, 2026 · arXiv:2608.19323

Improved Confidence Estimates for Black-Box Large Language Models

quantum machine learningquantum kernel methodspublished Aug 21, 2026 · arXiv:2608.19304

Quantum Kernel Estimation for the Discovery of Early Lung Cancer Detection

conversational AIintergroup relationspublished Aug 21, 2026 · arXiv:2608.19220

Can Conversational AI loosen Us-Versus-Them Boundaries? The Effects of Common, Dual, and Separate Identity Framings on Pro-Immigrant Intergroup Helping

LinkedGPT-5.6 Luna
supervised machine learninghelicopter weight estimationpublished Aug 21, 2026 · arXiv:2608.19210

Towards On-Board Implementation of ML-Based Helicopter Weight Estimator

language model self-improvementself-trainingpublished Aug 21, 2026 · arXiv:2608.20290

Phantom Gains: Auditing Self-Improvement Against a Measured Null

LinkedQwen 3.8 Max
cryptocurrency fraudmemecoinspublished Aug 21, 2026 · arXiv:2608.20271

Catching the Rug: Early Prediction of Fraudulent Memecoins on Solana via Machine Learning

quantum-classical neural networksphysical-layer authenticationpublished Aug 21, 2026 · arXiv:2608.20240

QUASAR: A Quantum-Classical Neural Network for SAR Satellite Physical-Layer Authentication

electronic navigational chartsgeospatial vector datapublished Aug 21, 2026 · arXiv:2608.20218

Electronic Navigational Chart Change Classification

Software 3.0software architecturepublished Aug 21, 2026 · arXiv:2608.20201

The Third Restructuring of Software Form: From the Three-Tier Architecture to Storage, Models, and Agents

compositional generalizationmulti-module language-model systemspublished Aug 21, 2026 · arXiv:2608.20054

What You Can't See Is What You Learn: Restricted Evidence Visibility Favors Compositional Generalization in Shared-Genome Language-Model Societies

AI agencymoral agencypublished Aug 21, 2026 · arXiv:2608.20041

A three-dimensional typology of agency for advanced AI systems

trajectory forecastingphysical property estimationpublished Aug 21, 2026 · arXiv:2608.20009

ExPhy: A Benchmark for Explicit Physical Property Learning in Multi-Object Trajectory Forecasting

time series forecastingTransformer architecturespublished Aug 21, 2026 · arXiv:2608.19966

Rethinking Patch Based Multivariate Time Series Forecasting with Semantic Structured Partitioning

mixed-integer linear programmingcombinatorial optimizationpublished Aug 21, 2026 · arXiv:2608.19953

Learning Early-to-Final Solution Consistency for MILP Acceleration

cardiac computed tomographystatistical shape modelspublished Aug 21, 2026 · arXiv:2608.19932

A Strong Linear Baseline for Whole-Heart Cardiac Shape Completion on CT, with an Open Eleven-Structure Statistical Shape Model

spiking neural networksBayesian inferencepublished Aug 21, 2026 · arXiv:2608.19907

Spike-based Belief Propagation in Nonlinear Dynamical Systems

LLM architecturedomain-specific languagespublished Aug 21, 2026 · arXiv:2608.19889

Write Once, Run Everywhere: The Axon DSL for Shape-Safe and Framework-Agnostic LLM Architectures

robustness testingadversarial robustnesspublished Aug 21, 2026 · arXiv:2608.19882

TESTNAV: Pareto-Guided Search for Compositional Robustness Testing

specification-driven developmentAI-assisted software developmentpublished Aug 21, 2026 · arXiv:2608.19838

Specification-delta-driven data governance: an empirical study of the "spec-delta" as the unit of change in lakehouse data platforms

AI content creationeducational technologypublished Aug 21, 2026 · arXiv:2608.19812

When Saying No Makes Better Videos: Designing Dual Gatekeeping for Pedagogically Grounded AI Content Creation

tensor networkstensor-train decompositionpublished Aug 21, 2026 · arXiv:2608.19789

TT-net: Quantum Inspired Tensor Network Denoising in Conditional GANs

ride-hailingorder dispatchingpublished Aug 21, 2026 · arXiv:2608.19751

GenMatch: An End-to-End Generative Matching Framework for Micro-View Order-Dispatching in Ride-Hailing

large language modelscontinual learningpublished Aug 21, 2026 · arXiv:2608.19680

Frequency-Aware Continual Learning for Smart Contract Vulnerability Detection with Large Language Models

AI controlAI safetypublished Aug 21, 2026 · arXiv:2608.19216

Bounded Sovereignty and the Control Tax: Pricing AI Oversight When the Deployer Does Not Own the Model

artificial consciousnessAI valencepublished Aug 21, 2026 · arXiv:2608.19215

How to Navigate Uncertainty About AI Consciousness

geographic biasinstitutional prestige biaspublished Aug 20, 2026 · arXiv:2608.18107

Institutional Prestige as Geographic Bias in Large Language Models: Evidence from Three Factorial Experiments with Bootstrap Confidence Intervals

low-resource languagesmultilingual language modelspublished Aug 20, 2026 · arXiv:2608.18094

NE-BERT: A Multilingual Language Model for Nine Northeast Indian Languages

model safetyabliterationpublished Aug 20, 2026 · arXiv:2608.18093

Abliteration Mitigation via Refusal Aliases

LinkedGemma 4
emotion representationvalence classificationpublished Aug 20, 2026 · arXiv:2608.18090

Nine Emotion Centroids: A Label-Free Valence Axis That Transfers Across Four Modalities

LinkedGemma 4Qwen 3.8 MaxMistral OCR 4
tokenizationnatural language processingpublished Aug 20, 2026 · arXiv:2608.18087

SuTRA : Structurally-Unified Tokenization with Root Awareness

frontier language modelsbenchmarkingpublished Aug 20, 2026 · arXiv:2608.19140

Grouping the Stochastic Machine: Precision, Not Capability, as the Frontier Metric for AI Systems

LLM operationsAI system reliabilitypublished Aug 20, 2026 · arXiv:2608.19125

Tuning the Stochastic Machine: A Systems Engineer's Operating Model for Human-AI Engineering

Wasserstein ambiguity setsentropic value-at-riskpublished Aug 20, 2026 · arXiv:2608.19073

Robust Risk Under Evolving Uncertainty: A Wasserstein Counterpart of the Entropic Value-at-Risk

self-promptingcross-model consensuspublished Aug 20, 2026 · arXiv:2608.19025

Self-prompting and cross-model consensus enable reproducible data extraction from scientific literature with large language models

deep learning testingmodel robustnesspublished Aug 20, 2026 · arXiv:2608.18900

TestifAI: Tomography-Based Testing for Deep Learning Systems

AI-assisted leak localizationverifiable abstentionpublished Aug 20, 2026 · arXiv:2608.18836

Verifiable abstention makes AI leak diagnosis accountable in water distribution networks

SARS-CoV-2variant detectionpublished Aug 20, 2026 · arXiv:2608.18238

GenEx: A Graph-Based Representational Paradigm for SARS-CoV-2 Variant Detection via Codon Co-occurrence Networks

lattice theorymetric spacespublished Aug 20, 2026 · arXiv:2608.18194

On the Triangle Inequality for the Jaccard Distance in Arbitrary Lattices

knowledge graphsRDFpublished Aug 20, 2026 · arXiv:2608.18165

RDFdL: Integrating RDF with Differential Dynamic Logic

Orientalismcultural sensitivitypublished Aug 20, 2026 · arXiv:2608.18100

Computational Orientalism: Measuring Structural Discourse Bias in Large Language Models Using the Middle East Cultural Sensitivity Score (MECSS)

LinkedGPT-5.6 Luna
Show more