Papers.
-
2026Skill Constellations: Tracing the Supply Chain of Agent Skills on GitHub
Builds the first dated copy network of 2.2M agent-skill adoptions on GitHub and uses it to rank which repositories to audit before risky skills spread.
-
2026 Under reviewPrincipled Thoughts for Latent Recursive LLM Systems
REST turns four properties of a valid thought (causality, minimality, separability, stability) into differentiable losses added to cross-entropy for latent single- and multi-agent LLM systems. With the same data and compute, it raises accuracy by up to 7.5 points across 7 benchmarks and adds no parameters at inference.
-
2026 ArabicNLP Arabic Natural Language Processing Conference @ EMNLP 2026 AcceptedFAAMG at AraSeg Shared Task 2026: Cross-Architecture Ensembling for Punctuation-Free Arabic Segmentation
A weighted ensemble of eight AraBERT, SaT, and CAMeLBERT BiLSTM-CRF models that segments Arabic text with no punctuation or paragraph breaks, fine-tuned on only 174 documents. It reaches 86.7 document-macro F1 on the blind test set and places 4th in AraSeg 2026 Subtask 4.
-
2026 FSE ACM International Conference on the Foundations of Software EngineeringPanther: Faster and Cheaper Computations with Randomized Numerical Linear Algebra
Randomized numerical linear algebra primitives that drop into existing PyTorch layers – 5x speedup over the dense baseline and 75% parameter reduction at negligible accuracy cost on BERT.
-
2026 Under reviewWhat Makes a Latent Thought Valid? Defining and Auditing Thought Representations for LLMs
Thought representations have four main properties to be valid.
-
2026 Under reviewARISE: A Repository-level Graph Representation and Toolset for Agentic Program Repair
Builds a multi-granularity repository graph with statement-level data-flow slicing so a program-repair agent can trace where a variable is defined and used in a single call. Mounted on SWE-agent, it resolves 4.7 more points of SWE-bench Lite issues and nearly doubles line-level fault-localization recall.
-
2026 TOSEM ACM Transactions on Software Engineering and Methodology Under reviewContext-Augmented Code Generation Using Programming Knowledge Graphs
Grounds code generation in a programming knowledge graph so the model retrieves precise structural context (call sites, types, idioms) before generating, instead of relying on lossy in-context recall.
-
2025 EMSE Empirical Software Engineering Under reviewAnalysis of AdvFusion: Adapter-based Multilingual Learning for Code LLMs
An empirical analysis of AdvFusion, an adapter-based multilingual training scheme for code LLMs, with a focus on cross-language transfer and pareto trade-offs.
-
2023 ICECCE International Conference on Electrical, Communication and Computer EngineeringAGS: Arabic GPT Summarization Corpus
An Arabic abstractive-summarisation dataset built via prompt engineering – the first Arabic abstractive corpus labelled by an LLM. Backed our 1st-place AIC-1 (ICMTC) submission.
3D model: "3D Origami crane" by JuanG3D · CC BY 4.0