Discussion about this post

User's avatar
Ian Tepoot's avatar

Thanks for this analysis. What you showed is a mathematical reality that many in our field don't want to confront. Your case that the adversarial subspace problem isn't a bug that can be patched, but a fundamental (and unmatchable) feature of how meaning gets flattened into numerical representations was convincing.

The brittleness you describe from the architectural mismatch of attempts to build human-like reasoning on foundations that can't support it was dead-on particularly if the focus is on the model level which is what the industry insists on tunnel-visioning. That to me suggests not just better guardrails (particularly in the way corporate tech views 'safety') but architecture where reasoning integrity itself is the 'resilience'.

Looking forward to hearing more...

No posts

Ready for more?