PREDATOR AI Research Dataset
50K+ AI/ML research papers from arXiv, NeurIPS, StackExchange
Curated AI research data including papers from arXiv, NeurIPS, ICML, and StackExchange discussions. Each entry includes source identifier, title, source platform, domain tags, value scores, and quality tier.
Fields
| Column | Description |
|---|---|
id |
ArXiv ID, paper DOI, or StackExchange question ID |
title |
Paper title or question summary |
source |
Platform (arXiv, NeurIPS, ICML, StackExchange) |
domain |
Classification (ai, nlp, cv, ml) |
commercial_value |
Commercial viability score (0-1) |
monetization_score |
Revenue potential score (0-2) |
quality |
Quality tier: high, medium, low |
Usage
import pandas as pd
df = pd.read_csv("data.csv")
nlp_data = df[df["domain"].str.contains("nlp|ai", case=False)]
print(f"NLP/AI entries: {len(nlp_data)}")
License
CC BY 4.0 โ Free for commercial use with attribution.
x402 API
curl -X POST "https://most-daylight-fraying.ngrok-free.dev/x402/ai/query" \
-H "X-PAYMENT: <tx_hash>"
$0.02 USDC/query โ Base chain, x402 protocol.
Contact
PREDATOR Data Platform โ iservice49800@gmail.com
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐ Ask for provider support