RELATIVITY ANALYTICS SPECIALIST EOC 2025/2026
QUESTIONS WITH ANSWERS GRADED A+
✔✔-Ensure server is in the workspaces Resource Pool
-Ensure the server has structured analytics operations box checked - ✔✔How would
you setup an analytics server to be available for selection in a workspace?
✔✔16 - ✔✔Regarding Clustering, if the maximum hierarchy depth is 1, what is the max
number of top level clusters that will be created?
✔✔-Disable optimize training set
-Full build - ✔✔What steps should you take to add documents back into the training set
after indexing?
✔✔They are at the beginning of a new line followed by a colon - ✔✔When are words
recognized as being part of an email header?
✔✔True - ✔✔T/F - The Go words filter is used with English Language Docs only
✔✔120 - 200 - ✔✔What is the normal range for average document size in words?
✔✔False - ✔✔T/F - Language ID uses a dictionary/word list?
✔✔White Space
Recipient meta
Conversation index
Email Action (RE, FW, FWD) - ✔✔What is not considered by the engine when marking
duplicates?
✔✔False - ✔✔T/F - An Email Duplicate Spare must also be an exact duplicate per MD5
or SHA256
✔✔Send
Reply
Replay-All
Forward
Draft - ✔✔What email Actions are identified by the analytics engine?
✔✔Message
Attachment
Inferred Match - ✔✔Name the reasons an email can be inclusive?
✔✔White Space
Capital Letters
, Punctuation
Numbers - ✔✔Name some things that Textual Near Duplicate ignores
✔✔.80 - 10.00 - ✔✔Consider Index Statistics: What is the normal range for unique
words per document?
✔✔If you change filters after the index has already been built - ✔✔When should you run
a full population?
✔✔True - ✔✔T/F - Email Threading only works on English Email Headers?
✔✔Coded due to meta
included due to family
too many concepts
bad text - ✔✔What makes a bad example document for categorization?
✔✔50 - ✔✔What is the default minimum coherence for a categorization set?
✔✔Message
Attachment
Inferred Match - ✔✔What are the inclusive reasons?
✔✔Word count
Punctuation marks
Number count
words with many characters - ✔✔What gets evaluated during optimize training set?
✔✔Each word in the index is compared to a valid word list on the server. If there is no
match for the word then it is ignored - ✔✔How does the Go Words filter work?
✔✔When they are at the beginning of a new line followed by a colon - ✔✔When are
words recognized as being part of an email header?
✔✔True - ✔✔T/F - The email header filter removed the word "Subject"
✔✔Outline and title - ✔✔What do Cluster names consist of?
✔✔Not Clustered - ✔✔What Cluster do documents that are not in the index go into?
✔✔Unclustered - ✔✔What Cluster do documents included in the index, but are without
searchable text go into?
QUESTIONS WITH ANSWERS GRADED A+
✔✔-Ensure server is in the workspaces Resource Pool
-Ensure the server has structured analytics operations box checked - ✔✔How would
you setup an analytics server to be available for selection in a workspace?
✔✔16 - ✔✔Regarding Clustering, if the maximum hierarchy depth is 1, what is the max
number of top level clusters that will be created?
✔✔-Disable optimize training set
-Full build - ✔✔What steps should you take to add documents back into the training set
after indexing?
✔✔They are at the beginning of a new line followed by a colon - ✔✔When are words
recognized as being part of an email header?
✔✔True - ✔✔T/F - The Go words filter is used with English Language Docs only
✔✔120 - 200 - ✔✔What is the normal range for average document size in words?
✔✔False - ✔✔T/F - Language ID uses a dictionary/word list?
✔✔White Space
Recipient meta
Conversation index
Email Action (RE, FW, FWD) - ✔✔What is not considered by the engine when marking
duplicates?
✔✔False - ✔✔T/F - An Email Duplicate Spare must also be an exact duplicate per MD5
or SHA256
✔✔Send
Reply
Replay-All
Forward
Draft - ✔✔What email Actions are identified by the analytics engine?
✔✔Message
Attachment
Inferred Match - ✔✔Name the reasons an email can be inclusive?
✔✔White Space
Capital Letters
, Punctuation
Numbers - ✔✔Name some things that Textual Near Duplicate ignores
✔✔.80 - 10.00 - ✔✔Consider Index Statistics: What is the normal range for unique
words per document?
✔✔If you change filters after the index has already been built - ✔✔When should you run
a full population?
✔✔True - ✔✔T/F - Email Threading only works on English Email Headers?
✔✔Coded due to meta
included due to family
too many concepts
bad text - ✔✔What makes a bad example document for categorization?
✔✔50 - ✔✔What is the default minimum coherence for a categorization set?
✔✔Message
Attachment
Inferred Match - ✔✔What are the inclusive reasons?
✔✔Word count
Punctuation marks
Number count
words with many characters - ✔✔What gets evaluated during optimize training set?
✔✔Each word in the index is compared to a valid word list on the server. If there is no
match for the word then it is ignored - ✔✔How does the Go Words filter work?
✔✔When they are at the beginning of a new line followed by a colon - ✔✔When are
words recognized as being part of an email header?
✔✔True - ✔✔T/F - The email header filter removed the word "Subject"
✔✔Outline and title - ✔✔What do Cluster names consist of?
✔✔Not Clustered - ✔✔What Cluster do documents that are not in the index go into?
✔✔Unclustered - ✔✔What Cluster do documents included in the index, but are without
searchable text go into?