| rdomains-package | rdomains: Classify Domains by their Content |
| adult_ml1_cat | Probability that Domain Hosts Adult Content Based on features of Domain Name and Suffix alone. |
| claude_cat | Get Category from Anthropic Claude |
| collect_content | Fetch homepage HTML and text for domains |
| dmoz_cat | Get Category from DMOZ |
| fetch_error_codes | Reasons a domain can fail to get a category |
| fetch_report | Summarise the outcome of a fetch or classification run |
| get_dmoz_data | Get DMOZ Data |
| get_shalla_data | Get Shalla Data |
| get_stevenblack_data | Get Steven Black's Host List Data |
| glm_shalla | ML Model |
| html_text_content | Extract text, title, description and language from HTML |
| not_news | Classify News and Non-News Based on keywords in the URL |
| openai_cat | Get Category from OpenAI |
| page_signals | What kind of page is this? |
| rdomains | rdomains: Classify Domains by their Content |
| shalla_cat | Get Category from Shallalist |
| source_vintage | Report what each category source is and when it was last published |
| stevenblack_cat | Get Category from Steven Black's Host List |
| uni_cat | Get Category from University Domain List |
| virustotal_cat | Get Category from VirusTotal |