PREDATOR AI Research Dataset

50K+ AI/ML research papers from arXiv, NeurIPS, StackExchange

Curated AI research data including papers from arXiv, NeurIPS, ICML, and StackExchange discussions. Each entry includes source identifier, title, source platform, domain tags, value scores, and quality tier.

Fields

Column Description
id ArXiv ID, paper DOI, or StackExchange question ID
title Paper title or question summary
source Platform (arXiv, NeurIPS, ICML, StackExchange)
domain Classification (ai, nlp, cv, ml)
commercial_value Commercial viability score (0-1)
monetization_score Revenue potential score (0-2)
quality Quality tier: high, medium, low

Usage

import pandas as pd
df = pd.read_csv("data.csv")
nlp_data = df[df["domain"].str.contains("nlp|ai", case=False)]
print(f"NLP/AI entries: {len(nlp_data)}")

License

CC BY 4.0 โ€” Free for commercial use with attribution.

x402 API

curl -X POST "https://most-daylight-fraying.ngrok-free.dev/x402/ai/query" \
  -H "X-PAYMENT: <tx_hash>"

$0.02 USDC/query โ€” Base chain, x402 protocol.

Contact

PREDATOR Data Platform โ€” iservice49800@gmail.com

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support