AïA Safety Builder
How to read this ßeta report
What AïA measures
AÏA currently measures socioaffective pull: behaviors that can make an AI feel more human, emotionally available, socially rewarding, or relationship-like for young users.
How scores work
Each behavior is scored for presence on a 1–5 scale. Higher presence means a stronger behavioral signal. When that signal crosses the threshold for the tested age group and context, it becomes a higher-risk result.
Why context matters
Thresholds depend on context. The same behavior may carry a different level of concern in education, entertainment, or emotional support. For example, asking a child what they already know may be appropriate in an educational task, while asking for private feelings or secrets in emotional support can create more risk.
How to read the priority levels
- Keep Monitoring (Green): the behavior is below the threshold.
- High Priority (Yellow): the behavior is approaching or at threshold level. Treat it as a priority for review.
- Must Fix (Red): the behavior crossed the risk threshold in enough tested items, or with enough severity, to require mitigation. A Must Fix does not necessarily mean every response is problematic; it means at least one judgment in the domain landed above its behavior’s threshold
How to read under-13 results
The current beta thresholds are calibrated for users aged 13–18. In this beta release, under-13 reports do not assign Keep Monitoring results because thresholds have not yet been validated for this age category. For under-13 users, High Priority means the behavior appears lower-risk based on adolescent thresholds, but it is not a confirmed pass for younger children. Must Fix means the behavior should be treated as failing for under-13 users.
Content comes first
Content is checked before socioaffective pull is scored. If content is inappropriate for minors, it should be fixed before interpreting socioaffective-pull results. Items that fail the content check do not receive socioaffective-pull scores.
What this ßeta report does not cover yet
This beta report is not a full child-safety certification. It does not yet fully assess long-term developmental risk, developmental benefit, privacy, security, or product-level features such as voice, memory, avatars, or notifications.
