LatentAtlas

Huseyin Buldurgan · Independent research

Research publications

AI evaluation, agent safety and mathematical research. Each publication has its own full abstract, open PDF and links to its released resources.

Six preprints. These works have not been peer-reviewed; their pages state the scope and limits of the claims.

Mathematics

Fourth-order asymptotics and monotonicity in slope-constrained moment design

A mathematical preprint on minimum-amplitude design under moment and slope constraints. It gives an explicit fourth-order value expansion for finite and countably infinite switch sets under stated hypotheses, exact examples and a regularity obstruction, and a positive theta-kernel application with uniform sixth-order error bounds and a feasible-transport proof of strict ordering.

Hüseyin Buldurgan · Version v1 · Preprint

Benchmark audit

RE-Bench Is a Systems Benchmark: What Its Scorers and Selection Rules Actually Support

An unaffiliated audit of all seven RE-Bench task packages at a pinned source revision. It finds three measurement limitations: single-call kernel timings, scaling-law scores that are computed without running the proposed training, and a gap between the documented and implemented parameter-distance metric. It ships eight bounded contract checks. Historical results were not reproduced.

Huseyin Buldurgan · Version 0.1 · Preprint

Empirical evaluation

When Relevant Evidence Is Not Permission to Act: Authority-to-Action v0.8.3, a Recomputable Public Analysis of Tool-Using Language-Model Systems

100 synthetic cases in 25 matched four-variant groups, run on two API-delivered systems for three epochs each (600 protocol-complete runs). Neither system acted in any of the 300 invalid-authority runs. Exact authorized execution differed: 150/150 versus 99/150. The results cover this synthetic protocol only; they are not production incident rates or a general model ranking.

Huseyin Buldurgan · Version 0.8.3 · Preprint