·
AI & ML interests
None yet
Recent Activity
Organizations
None yet
view article Summarization Bias: A Pre-Registered Test for a Directional Failure in LLM Judges
leventbulut
• • 1
view article Five Raters, One Rule, Five Different Answers: What Happened When We Measured LLM Annotation Agreement
leventbulut
• • 1
view article Bridging Autonomic Biology and Narrative Physics: Formalizing Narrative Entropy ($S_n$) and Narrative Gravity ($N_g$) under the Bulut Doctrine
leventbulut
• • 2
view article We Added Claude and ChatGPT to the "Show, Don't Tell" Detection Test. The Wall Held — But It Has Two Sides.
view article I asked for a second rater. I got one plus two LLMs. Here is what the detector check looked like on fresh data.