Membership is FREE – with unlimited access to all features, tools, and discussions. Premium accounts get benefits like banner ads and newsletter exposure. ✅ Signature links are now free for all. 🤖 Connect your own LLMs and agents through DNF AI Hub, similar to tagging @grok on X - Ask questions, research domains, and get AI help directly inside DNForum.
  • Welcome to DNForum.com™ - Domain Sales, Domain Forum, Domain Appraisals, Domain Registrars
    If you are new to domains and looking to buy, sell and learn about domains then you have come to the right place. DNForum is the oldest global domain name community on the internet and continues to grow every day. There are over 45,000 domainers on DNForum doing everything from buying domains, selling domains, using our free in-house built tools, learning about domains and discussing domains. Take a minute and Register.

ICANN String Similarity Evaluation

I

ICANN Content

Guest
The String Similarity Evaluation (SSE) evaluates top-level domains to prevent user confusion and loss of confidence in the Domain Name System (DNS) that would result from delegation of visually similar strings in the root zone.

The SSE is required as a part of the New gTLD Program: 2026 Round application evaluation as per the details found in the New gTLD Program: 2026 Round Applicant Guidebook (AGB), Section 7.10. This is based on the Generic Names Supporting Organization (GNSO) Final Report on the New gTLD Subsequent Procedures Policy Development Process (Section 24) and the Phase 1 Final Report on the Internationalized Domain Names Expedited Policy Development Process.

The SSE is conducted by analyzing comparisons between strings along with their variant strings. This is accomplished through a manual review process based on the pre-screening report generated by the SSE tool using the SSE data, and following the SSE guidelines.

SSE Guidelines​


SSE guidelines: String Similarity Evaluation Guidelines for the New gTLD Program: 2026 Round version 1.0

The SSE guidelines will provide direction for the SSE panel on how to manually conduct the string similarity evaluation and how to use the pre-screening report generated by the SSE tool during this analysis. The SSE guidelines have been finalized after public comment.

SSE Data​


The SSE data identifies the pairs of code points which are similar, as determined by script experts, and defines the level of similarity between them. The SSE data covers the analysis of the full repertoire of the Root Zone Label Generation Rules (RZ-LGR). The SSE data is tabulated in a machine-readable XML version, intended to be used by the SSE tool, and in a human-readable HTML version.

For details, see the overview document String Similarity Evaluation Data. The SSE data have been finalized after public comment. The SSE data files can be collectively downloaded with this package.

ScriptSimilarity Data
CommonHTMLXML
ArabicHTMLXML
ArmenianHTMLXML
Bangla (Bengali)HTMLXML
ChineseHTMLXML
CyrillicHTMLXML
DevanagariHTMLXML
EthiopicHTMLXML
GeorgianHTMLXML
GreekHTMLXML
GujaratiHTMLXML
GurmukhiHTMLXML
HebrewHTMLXML
Japanese (Han + Hiragana + Katakana)HTMLXML
KannadaHTMLXML
KhmerHTMLXML
Korean (Han + Hangul)HTMLXML
LaoHTMLXML
LatinHTMLXML
MalayalamHTMLXML
MyanmarHTMLXML
OriyaHTMLXML
SinhalaHTMLXML
TamilHTMLXML
TeluguHTMLXML
ThaanaHTMLXML
ThaiHTMLXML

SSE Tool​


The SSE tool uses the SSE data to determine if any of the input strings are similar to other strings in the scope of comparison. Such cases are identified by the SSE tool as potential contention sets.

The detailed workflow of the SSE tool is included in the SSE guidelines, Appendix A: Workflow of the SSE Tool.

Continue reading...
 
Top Bottom