Discourse AI의 독성 감지, 다음 단계는 무엇일까요?

오늘 우리는 Discourse AI - Toxicity 모듈에 대한 작별을 발표합니다 :waving_hand:. 대신 https://meta.discourse.org/t/discourse-ai-post-classifier-automation-rule/281227를 활용하여 대규모 언어 모델(LLM)의 힘을 빌려 더 나은 경험을 제공하고자 합니다.

왜 이런 조치를 취하는가?

이전에는 Toxicity 모듈을 사용하면 다음과 같은 한계가 있었습니다.

  • 미리 정의된 단일 모델만 사용해야 했습니다.
  • 커뮤니티의 특정 필요에 따른 커스터마이징이 불가능했습니다.
  • 혼란스러운 임계값 지표가 존재했습니다.
  • 성능이 기대에 미치지 못했습니다.

LLM은 크게 발전하여 이제 더 나은 성능과 커스터마이징 가능한 경험을 제공할 수 있게 되었습니다.

무엇이 새로운가?

https://meta.discourse.org/t/discourse-ai-post-classifier-automation-rule/281227는 Toxicity(독성) 게시물 분류(그 밖의 여러 용도 포함)를 처리하고 커뮤니티가 특정 행동 강령을 따르도록 강제하는 데 사용할 수 있습니다. 이는 다음과 같은 것을 의미합니다.

  • 다양한 성능 요구 사항에 대응하기 위해 여러 LLM 지원
  • 콘텐츠가 어떻게 처리되어야 하는지 쉽게 정의 가능
  • 커뮤니티의 특정 필요에 맞는 프롬프트 커스터마이징
  • 검토를 위해 콘텐츠 플래깅

그리고 그 밖에도 더 많은 기능이 있습니다.

전이를 돕기 위해 이미 가이드를 작성해 두었습니다.

Toxicity는 어떻게 되는가?

이 발표는 아주 초기 단계로 간주되어야 합니다. 폐기 준비가 될 때까지 Toxicity를 계속 사용할 수 있습니다. 폐기 시점에는 해당 모듈을 폐기하고 Discourse AI 플러그인 및 관련 서비스에서 모든 코드를 제거할 것입니다.

:backhand_index_pointing_right:t5: 업데이트: Toxicity 모듈은 이제 공식적으로 Discourse에서 제거되었습니다. 이는 모든 관련 사이트 설정 및 기능을 포함합니다. 이제 사용자에게 Discourse AI - AI 분류로 전환하고 위에서 나열된 가이드를 따를 것을 권고합니다.

Business 및 Enterprise 고객은 사이트의 관리자 설정에서 What's New 섹션 아래에 다음 내용을 확인하게 됩니다. 이를 통해 추가 비용 없이 Discourse가 호스팅하는 LLM을 사용하여 AI 분류를 활성화할 수 있습니다.

6개의 좋아요

Is there a cut-off date for the Toxicity module?

3개의 좋아요

We haven’t picked a specific date but are currently working on getting this done soon

3개의 좋아요

This is extremely relevant to an article I’m writing for Community Leaders Institute right now, to my Ph.D on toxicity, and more. Would you all be willing to do a behind-the-scenes interview on this system for my YouTube channel as part of my series on inoculating online communities from toxicity? It is also doubling as an academic endeavor for my Ph.D.

You up for that?

2개의 좋아요

Thank you for thinking about Discourse for your YouTube channel! Could you please email more details to mae@discourse.org about what exactly the behind the scenes interview would entail and what’s needed from our end?

2개의 좋아요

FYI about the following change

2개의 좋아요

Thanks Mae! I’ve followed up!

1개의 좋아요

As a heads-up, we are now hiding site settings for enabling/disabling Toxicity and NSFW. This is our ongoing effort as we continue to deprecate the features.

If you have these features turned on, it will still operate as normal. We have NOT fully deprecated the features yet.

If you have it turned off and would like to turn it on, you will now be unable to do so.

1개의 좋아요

여러분, 업데이트 내용을 간단히 공유드리고자 합니다

1개의 좋아요