·
AI & ML interests
None yet
Recent Activity
Organizations
None yet
view article Summarization Bias: A Pre-Registered Test for a Directional Failure in LLM Judges
leventbulut
• • 1
view article Five Raters, One Rule, Five Different Answers: What Happened When We Measured LLM Annotation Agreement
leventbulut
• • 1
view article Bridging Autonomic Biology and Narrative Physics: Formalizing Narrative Entropy ($S_n$) and Narrative Gravity ($N_g$) under the Bulut Doctrine
leventbulut
• • 2
published an article about 1 month ago view article I named my framework "physics." Then I checked whether it survives the word.
published an article about 2 months ago view article Five machine raters, one definition, and answers ranging from 0 to 78
view article We Added Claude and ChatGPT to the "Show, Don't Tell" Detection Test. The Wall Held — But It Has Two Sides.
view article I asked for a second rater. I got one plus two LLMs. Here is what the detector check looked like on fresh data.
view article What happens when you check whether your own detector detects what it claims
view article Objective Projection v7.2: Closing the Pattern F Gap and Extending the Hard Negatives
view article Summarization Bias: Why Language Models Re-Label the Emotions You Tried to Hide
view article Objective Projection: Engineering Emotion in Text with Physical Parameters Instead of Emotion Labels