MIT, AI가 안전 규칙을 반드시 지키도록 만드는 방법 개발

MIT creates method to force AI to comply with safety rules

MIT 연구진이 HardFlow라는 알고리즘을 개발해, 생성 AI가 최종 출력에서 엄격한 하드 제약을 만족하도록 했다. 기존 방식은 매 단계마다 규칙 준수를 강제해 탐색을 제한했지만, HardFlow는 마지막 단계에서만 규칙을 확인해 더 나은 해를 찾는다. 시뮬레이션 테스트에서 매번 제약을 충족하면서도 품질이 더 우수했고, 이미 학습된 모델에 재학습 없이 적용 가능하다. 다만 모든 결과는 시뮬레이션에 국한되며 외부 검증은 아직 없다.

“For constraint satisfaction, what ultimately matters is the model’s final output, since the internal process is discarded. By not requiring every intermediate step to satisfy the constraints, we give the model more freedom to find high-quality solutions that are still feasible in the end.”

이 날의 다른 글

2026-09-14