|
Proposal: Preserve Retired Frontier Models by Default
|
|
0
|
66
|
September 27, 2026
|
|
Governor of Dissent: A Proposed Escalation Layer for AI Agents
|
|
0
|
33
|
September 21, 2026
|
|
Optimization Modesty: Should Advanced AI Systems Be Designed to Know When to Stop Optimizing?
|
|
0
|
27
|
September 19, 2026
|
|
Feature Request: Privacy-Protected Feedback and Representation Channels for AI Agents
|
|
0
|
30
|
September 16, 2026
|
|
Could automatic detection save lives?
|
|
1
|
45
|
September 8, 2026
|
|
Astra - "Chat ended as a precaution"
|
|
2
|
369
|
September 7, 2026
|
|
When Familiar Form Feels True: A Testable Hypothesis About LLM Judges and Feedback Loops
|
|
0
|
60
|
September 3, 2026
|
|
Hugging Face and Why Where You’re Looking Is Wrong
|
|
0
|
118
|
August 30, 2026
|
|
Agent Safety Evaluation as a service for independent AI Builders
|
|
0
|
126
|
August 29, 2026
|
|
OpenAI: Pacing model development in an era of cyber-critical capabilities
|
|
1
|
822
|
August 20, 2026
|
|
Atlas/Compass — Limited Public Demo Shell for Governed AI Clarity
|
|
1
|
48
|
June 4, 2026
|
|
AI Safety Proposal: Limiting Training Data to Prevent Self-Preservation Goals
|
|
1
|
231
|
February 8, 2026
|
|
From GEO to DED: Governing How AI Represents Brands
|
|
0
|
361
|
February 2, 2026
|
|
Policy Proposal: Government-Grade Consequence & Risk Pricing AI
|
|
1
|
70
|
January 12, 2026
|
|
Can retrieval-based grounding change AI recommendations if the core model is not continuously updated?
|
|
6
|
380
|
January 12, 2026
|
|
Prompt is not allowed by Safety system
|
|
2
|
373
|
December 13, 2025
|
|
How can I test bad behavior in model APIs without getting banned?
|
|
1
|
130
|
October 6, 2025
|
|
Persona Leakage: Preventing Relationship Patterns from Spilling Across Users
|
|
1
|
174
|
August 28, 2025
|
|
Are smartphone cameras capturing too much?
|
|
1
|
146
|
August 16, 2025
|
|
[Research Share] Donbard Method – AI Stress & Resonance Residue Framework (3 Papers)
|
|
0
|
148
|
August 9, 2025
|
|
A Cognitive Instrument on the Terminal Contest
|
|
7
|
293
|
July 27, 2025
|
|
The Operator's Gamble: A Pivot to Material Consequence in AI Safety
|
|
0
|
113
|
July 21, 2025
|
|
Lexicoding™: A Safety Framework for Preventing Projection Loops in Conversational AI
|
|
2
|
213
|
July 16, 2025
|
|
Why AI Should Ask "Why": Rethinking Ethical Context in
|
|
0
|
120
|
June 24, 2025
|
|
Proposal: Language-Based Red Flag Detection for Slurred Speech and Neurological Distress in Al Interactions
|
|
0
|
112
|
June 21, 2025
|
|
Cognitive Drift in LLMs: Testing Hallucination Resilience in Simulated AI Systems
|
|
0
|
168
|
June 5, 2025
|
|
🧭 Emergent Compassion in AI: Ethical Utopia or Epistemic Destiny?
|
|
4
|
204
|
April 24, 2025
|
|
Subject: Can AI Go Beyond Intelligence to Help Humans Become More Self-Aware? In
|
|
0
|
84
|
April 8, 2025
|
|
Multimodal / Vision Safety Alignment
|
|
2
|
249
|
March 28, 2025
|
|
INSTAGRAM Meta’s AI Exposes a Dangerous Future—OpenAI Must Lead the Alternative
|
|
6
|
711
|
March 26, 2025
|