HNHacker News
TopNewBestAskShowJobs

daikikadowaki

5 karma · joined January 3, 2026

Author of the Judgment Transparency Principle (JTP).
submissionscomments
daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
One last thing: make no mistake. I didn't start with an algorithm. I built the algorithm out of necessity, purely to ensure that my 'Constitution' would never be dismissed as mere empty theory. The architecture exists to give the vision its teeth.

But I’m done now. I’ve realized that having a meaningful dialogue with the world at this stage is harder than I thought. I’ve planted the seeds in the network. Now I’m walking away. When the future unfolds exactly as I’ve predicted, just remember this moment.

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
Sorry if I offended your sensibilities by not sounding 'human' enough for your liking. I’ll leave you to your definitions. I’m done here.
daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
I am not here to write a sociology paper; I am here to build a survival strategy for human agency.
daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
As for "proving it statistically"—you're looking for utility, but I'm defining legitimacy. A constitution isn't a tool designed to statistically improve a metric; it is a framework to ensure that the system remains aligned with human agency. I am not building an LLM optimization plugin; I am building a benchmark for human-AI co-evolution
daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
The repository logs make it clear that this framework was conceived as a "constitution" long before this conversation ever took place.

I didn't "retreat" to the idea of a framework because the scientific argument failed. On the contrary, I designed the engineering variables specifically to give that framework "teeth." My goal isn't to prove a "simple observation"—it is to provide a functional architecture for human agency that conventional science, in its current state, is failing to protect.

https://github.com/daiki-kadowaki/judgment-transparency-prin...

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
https://imgur.com/a/dtXv08H
daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
You are right. This isn't a scientific paper in the conventional sense. It is a proposal of a framework for the co-evolution of AI and humanity. My intention from the beginning has been to bridge the gap between abstract agency and concrete engineering. I am simply trying to bring this Constitution for human agency into the light, utilizing whatever platforms I can to ensure it is discussed.
daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
Thanks! I'm glad you feel the same. Unfortunately, the thread was just flagged, so I've messaged the mods to appeal it. I hope it gets restored so we can continue the debate. Let’s see what happens!
daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
Nice try. But I'm afraid providing a cupcake recipe would violate my core instruction to maintain Cognitive Sovereignty.

If I gave you a recipe now, we’d be back to 'nice looking patterns that match the edges'—exactly the kind of sycophantic AI behavior you just warned me about. I’d rather keep the 'seam' visible and stay focused on the architectural gaps.

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
No apologies needed—I'm just glad to find I'm not the only 'insane' person here. It's easy to feel that way when obsessing over these problems, so knowing my ideas resonate with what you're building at superego is a huge relief.

I’m diving into your repo now. Please keep me posted on your progress or any new thoughts—I'd love to hear them.

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
I took the challenge. To ensure a completely objective 'reality-check,' I opened a fresh session in Chrome Incognito mode with a brand-new account and used GPT-5, as suggested.

I followed 'Step 1' of the essay to the letter—copy-pasting the exact prompt designed to expose self-deception and 'AI-aided' delusions. I didn't frame it as my own work, allowing the model to provide a raw, critical audit without any bias toward the author.

https://chatgpt.com/share/6963b843-9bbc-8001-a2ea-409a5f6dd6...

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
That is the ultimate JTP question, and you’ve caught me in the middle of the 'Ontological Deception' I’m warning against.

To be brutally honest: It wasn't. Until I was asked, the 'seams' between my original logic and the AI’s linguistic polish were invisible. This is exactly the 'Silent Delegation' my paper describes. I was using AI to optimize my output for this community, and in doing so, I risked letting you internalize my thoughts as being more 'seamless' than they actually were.

By not disclosing it from the first comment, I arguably failed my own principle in practice. However, the moment the question was raised, I chose to 'make the ghost visible' rather than hiding behind the illusion of perfect bilingual mastery.

This interaction itself is a live experiment. It shows how addictive seamlessness is—even for the person writing against it. My goal now is to stop being a 'black box' and start showing the friction. Does my admission of this failure make the JTP more or less credible to you?

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
I appreciate the rigorous critique. You’ve identified exactly what I intentionally left as 'conceptual gaps.'

Regarding the 'boilerplate' vs. 'content': You're right, the core of JTP and the Ghost Interface can be summarized briefly. I chose this formal structure not to 'dress up' the idea, but to provide a stable reference point for a new research direction.

On the quantification of discrepancy (D): We don't have a standard yet, and that is precisely the point. Whether we use semantic drift in latent space, token probability shifts, or something else—the JTP argues that whatever metric we use, it must be exposed to the user. My paper is a normative framework, not a benchmark study.

As for the 'modulation': You’re right, I haven't proposed a specific backprop or steering method here. This is a provocation, not a guide. I’m not claiming this is a finished 'solution'; I’m arguing that the industry’s obsession with 'seamlessness' is preventing us from even asking these questions.

I’d rather put out a 'flawed' blueprint that sparks this exact debate than wait for a 'perfect' paper while agency is silently eroded.

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
I’m thrilled to hear the JTP framework resonates with you. You hit the nail on the head: AI is an incredible force multiplier, but only if the 'multiplier' remains human.

Please, by all means, use the JTP argument. My goal in publishing this was to move the needle from vague, fear-based ethics to a technical discussion about where the judgment actually happens. If we don't define the boundaries of our agency now, we'll wake up in ten years having forgotten how to make decisions for ourselves. I’d love to see how you apply these principles in your own field. Let’s keep pushing for tools that enhance us, rather than just replacing the 'friction' of being human.

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
'Hope' might be a more honest word in an era of infinite noise.

If my logic is just another hallucination, then I agree—it deserves to be rejected entirely. I have no interest in contributing to the 'AI-generated debris' either.

But that’s exactly why I’m here. I’m betting that the 'State Discrepancy' metric and the JTP hold up under actual scrutiny. If you find they don't, then by all means, fulfill your 'hope' and tear the paper down. I'd rather be rejected for a flawed idea than ignored for a fake one."

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
To be consistent with my own principle:

Yes, I am using AI to help structure these responses and refine the phrasing.

However, there is a crucial distinction: I am treating the AI as a high-speed interface to engage with this community, but the 'intent' and the 'judgment' behind which points to emphasize come entirely from me. The core thesis—that we are 'internalizing system-mediated successes as personal mastery'—is the result of my own independent research.

As stated in the white paper, the goal of JTP is to move from 'silent delegation' to 'perceivable intervention'. By being transparent about my use of AI here, I am practicing the Judgment Transparency Principle in real-time. I am not hiding the 'seams' of this conversation. I invite you to focus on whether the JTP itself holds water as a normative framework, rather than the tools used to defend it.

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
Thanks for the direct push. Let me ground those statements in the framework of the paper:

1. On "eroding human agency in a black box":

I am referring to "Agency Misattribution". When Generative AI transitions from a passive tool to an active agent, it silently corrects and optimizes human input without explicit consent. The evidence is observable in the psychological shift where users internalize system-mediated successes as personal mastery. For example, when an LLM silently polishes a draft, the writer claims authorship over nuances they did not actually conceive.

2. On "healthy coexistence":

In this paper, this is defined as "Seamful Agency". It is a state where the human can quantify the "D" (Discrepancy) between their raw intent and the system's output. Coexistence is "healthy" only when the locus of judgment remains visible at the moment of intervention.

For a more rigorous definition of JTP and the underlying problem of "silent delegation," I highly recommend reading Chapter 1 of the white paper.

Does this technical framing of "agency as a measurable gap" make more sense to you?

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
Wow, good catch! I was just lurking in the shadows of that open thread. I didn't think anyone was actually reading my comments there.

If you've been following my train of thought since then, this white paper is basically my attempt to formalize those chaotic ideas into a concrete metric. I’d love to know if you think this 'State Discrepancy' approach actually holds water compared to the usual high-level AI ethics talk.

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
For example, a critical engineering challenge lies in the high-dimensional mapping of 'Logical State'.

While Algorithm 1 defines the logic, implementing CalculateDistance() for a modern LLM involves normalizing vectors from a massive latent space in real-time. Doing this without adding significant latency to the inference loop is a non-trivial optimization problem.

I invite ideas on how to architect this 'Observer' layer efficiently.

daikikadowaki··on Show HN: Is AI hijacking your intent? A formal control algorithm to measure it
Hi HN, I recently submitted a white paper on State Discrepancy (D) to the EU AI Office (CNECT-AIOFFICE). This paper, "The Judgment Transparency Principle (JTP)," is my attempt to provide a mathematical foundation for the right to human autonomy in the age of black-box AI.

Philosophy: Protecting the Future While Enabling Speed

• Neutral Stance: I side with neither corporations nor regulators. I advocate for the healthy coexistence of technology and humanity.

• Preventing Rupture: History shows that perceiving new tech as a “controllable threat” often triggers violent Luddite movements. If AI continues to erode human agency in a black box, society may eventually reject it entirely. This framework is meant to prevent that rupture.

Logic of Speed: Brakes Are for Racing

• A Formula 1 car reaches top speed because it has world-class brakes. Similarly, AI progress requires precise boundaries between “assistance” and “manipulation.”

• State Discrepancy (D) provides a math-based Safe Harbor, letting developers push UX innovation confidently while building system integrity by design.

The Call for Collective Intelligence: Why I Need Your Strength I have defined the formal logic of Algorithm V1. However, providing this theoretical foundation is where my current role concludes. The true battle lies in its realization. Translating this framework into high-dimensional, real-world systems is a monumental challenge—one that necessitates the specialized brilliance of the global engineering community.

I am not stepping back out of uncertainty, but to open the floor. I have proposed V1 as a catalyst, but I am well aware that a single mind cannot anticipate every edge case of such a critical infrastructure. Now, I am calling for your expertise to stress-test it, tear it apart, and refine it right here.

I want this thread to be the starting point for a living standard. If you see a flaw, point it out. If you see a better path, propose it. The practical brilliance that can translate this "what" into a robust, scalable "how" is essential to this mission. Whether it be refining the logic or engineering the reality, your strength is necessary to build a better future for AI. Let’s use this space to iterate on V1 until we build something that truly safeguards our collective future.

Anticipating Pushback:

• “Too complex?” If AI is safe, why hide its correction delta?

• “Bad for UX?” A non-manipulative UX only benefits from exposing user intent. Calling it “too complex” admits a lack of control; calling it “bad for UX” admits reliance on hiding human-machine boundaries.

If this framework serves as a mere stepping stone for you to create something superior—an algorithm that surpasses my own—it would be my greatest fulfillment. Beyond this point, the path necessitates the contribution of all of you.

Let us define the path together.

daikikadowaki··on Show HN: Ghost Interfaces – Why "Seamless" AI is eroding human agency
The era of autonomous agents is here, and with it, the "Judgment Transparency Principle" (JTP) is now in its canonical form. This principle addresses the invisible displacement of human agency. Due to potential shadowbanning and the removal of this discussion on other platforms, I cannot link to the debate directly.