XenArcAI

company

XenArcAI

xenarcai

AI & ML interests

At XenArcAI, we don't just build with AI ,we understand it. Born at the edge of research and driven by a hunger to innovate, our mission is simple yet powerful: Turn cutting-edge science into real-world intelligence. We believe innovation without understanding is noise. And research without impact is dust. That’s why we exist — to bridge the lab and the world. Our team combines deep learning with deep thinking, transforming complexity into clarity, and theory into tools that move industries. We’re not chasing trends. We’re building what’s next. 🔬 Research-Rooted 🚀 Innovation-Driven

Recent Activity

Parveshiiii updated a model about 7 hours ago

XenArcAI/AIRealNet

Parveshiiii updated a model about 7 hours ago

XenArcAI/SparkEmbedding-300m

Parveshiiii updated a dataset 9 days ago

XenArcAI/MathX-5M

View all activity

Parveshiiii

updated 2 models about 7 hours ago

XenArcAI/AIRealNet

Image Classification • 0.2B • Updated about 7 hours ago • 4.41k • • 4

XenArcAI/SparkEmbedding-300m

Parveshiiii

updated a dataset 9 days ago

XenArcAI/MathX-5M

Viewer • Updated 9 days ago • 4.32M • 11.6k • 65

Parveshiiii

updated a Space 10 days ago

README

🚀

Parveshiiii

posted an update 13 days ago

Post

1587

Another banger from XenArcAI! 🔥

We’re thrilled to unveil three powerful new releases that push the boundaries of AI research and development:

🔗 XenArcAI/SparkEmbedding-300m

- A lightning-fast embedding model built for scale.
- Optimized for semantic search, clustering, and representation learning.

🔗 XenArcAI/CodeX-7M-Non-Thinking

- A massive dataset of 7 million code samples.
- Designed for training models on raw coding patterns without reasoning layers.

🔗 XenArcAI/CodeX-2M-Thinking

- A curated dataset of 2 million code samples.
- Focused on reasoning-driven coding tasks, enabling smarter AI coding assistants.

Together, these projects represent a leap forward in building smarter, faster, and more capable AI systems.

💡 Innovation meets dedication.
🌍 Knowledge meets responsibility.

Parveshiiii

in XenArcAI/CodeX-2M-Thinking 13 days ago

[bot] Conversion to Parquet

#1 opened 14 days ago by

parquet-converter

Parveshiiii

updated 2 datasets 14 days ago

XenArcAI/CodeX-2M-Thinking

Viewer • Updated 14 days ago • 2.19M • 1.71k • 6

XenArcAI/CodeX-7M-Non-Thinking

Viewer • Updated 14 days ago • 7.36M • 1.77k • 6

Parveshiiii

updated a collection 15 days ago

CodeX

Collection

The best available Pre-curated coding datasets on Platform. • 3 items • Updated 15 days ago

Parveshiiii

posted an update 20 days ago

Post

3008

SparkEmbedding - SoTA cross lingual retrieval

Iam very happy to announce our latest embedding model sparkembedding-300m base on embeddinggemma-300m we fine tuned it on 1m extra examples spanning over 119 languages and result is this model achieves exceptional cross lingual retrieval

Model: XenArcAI/SparkEmbedding-300m

Parveshiiii

published a model 23 days ago

XenArcAI/SparkEmbedding-300m

Parveshiiii

updated a collection 23 days ago

CodeX

Collection

The best available Pre-curated coding datasets on Platform. • 3 items • Updated 15 days ago

Parveshiiii

posted an update about 2 months ago

Post

193

AIRealNet - SoTA - Image detection model

We’re proud to release AIRealNet — a binary image classifier built to detect whether an image is AI-generated or a real human photograph. Based on SwinV2 and fine-tuned on the AI-vs-Real dataset, this model is optimized for high-accuracy classification across diverse visual domains.

If you care about synthetic media detection or want to explore the frontier of AI vs human realism, we’d love your support. Please like the model and try it out. Every download helps us improve and expand future versions.

Model page: XenArcAI/AIRealNet

Parveshiiii

posted an update about 2 months ago

Post

4480

Ever wanted an open‑source deep research agent? Meet Deepresearch‑Agent 🔍🤖

1. Multi‑step reasoning: Reflects between steps, fills gaps, iterates until evidence is solid.

2. Research‑augmented: Generates queries, searches, synthesizes, and cites sources.

3. Fullstack + LLM‑friendly: React/Tailwind frontend, LangGraph/FastAPI backend; works with OpenAI/Gemini.

🔗 GitHub: https://github.com/Parveshiiii/Deepresearch-Agent

Parveshiiii

posted an update 2 months ago

Post

3098

🚀 Big news from XenArcAI!

We’ve just released our new dataset: **Bhagwat‑Gita‑Infinity** 🌸📖

✨ What’s inside:
- Verse‑aligned Sanskrit, Hindi, and English
- Clean, structured, and ready for ML/AI projects
- Perfect for research, education, and open‑source exploration

🔗 Hugging Face: XenArcAI/Bhagwat-Gita-Infinity

Let’s bring timeless wisdom into modern AI together 🙌

Parveshiiii

posted an update 2 months ago

Post

2450

🚀 New Release from XenArcAI
We’re excited to introduce AIRealNet — our SwinV2‑based image classifier built to distinguish between artificial and real images.

✨ Highlights:
- Backbone: SwinV2
- Input size: 256×256
- Labels: artificial vs. real
- Performance: Accuracy 0.999 | F1 0.999 | Val Loss 0.0063

This model is now live on Hugging Face:
👉 XenArcAI/AIRealNet

We built AIRealNet to push forward open‑source tools for authenticity detection, and we can’t wait to see how the community uses it.

Parveshiiii

posted an update 4 months ago

Post

1109

🚀 Just Dropped: MathX-5M — Your Gateway to Math-Savvy GPTs

👨‍🔬 Wanna fine-tune your own GPT for math?
🧠 Building a reasoning agent that actually *thinks*?
📊 Benchmarking multi-step logic across domains?

Say hello to [**MathX-5M**]( XenArcAI/MathX-5M) — a **5 million+ sample** dataset crafted for training and evaluating math reasoning models at scale.

Built by **XenArcAI**, it’s optimized for:
- 🔍 Step-by-step reasoning with , , and formats
- 🧮 Coverage from arithmetic to advanced algebra and geometry
- 🧰 Plug-and-play with Gemma, Qwen, Mistral, and other open LLMs
- 🧵 Compatible with Harmony, Alpaca, and OpenChat-style instruction formats

Whether you're prototyping a math tutor, testing agentic workflows, or just want your GPT to solve equations like a pro—**MathX-5M is your launchpad**.

🔗 Dive in: ( XenArcAI/MathX-5M)

Let’s make open-source models *actually* smart at math.
#FineTuneYourGPT #MathX5M #OpenSourceAI #LLM #XenArcAI #Reasoning #Gemma #Qwen #Mistral

Parveshiiii

posted an update 4 months ago

Post

1095

🚀 Launch Alert: Dev-Stack-Agents
Meet your 50-agent senior AI team — principal-level experts in engineering, AI, DevOps, security, product, and more — all bundled into one modular repo.

+ Code. Optimize. Scale. Secure.
- Full-stack execution, Claude-powered. No human bottlenecks.

🔧 Built for Claude Code
Seamlessly plug into Claude’s dev environment:

* 🧠 Each .md file = a fully defined expert persona
* ⚙️ Claude indexes them as agents with roles, skills & strategy
* 🤖 You chat → Claude auto-routes to the right agent(s)
* ✍️ Want precision? Just call @agent-name directly
* 👥 Complex task? Mention multiple agents for team execution

Examples:

"@security-auditor please review auth flow for risks"
"@cloud-architect + @devops-troubleshooter → design a resilient multi-region setup"
"@ai-engineer + @legal-advisor → build a privacy-safe RAG pipeline"

🔗 https://github.com/Parveshiiii/Dev-Stack-Agents
MIT License | Claude-Ready | PRs Welcome

Parveshiiii

posted an update 5 months ago

Post

2704

🧠 Glimpses of AGI — A Vision for All Humanity
What if AGI wasn’t just a distant dream—but a blueprint already unfolding?

I’ve just published a deep dive called Glimpses of AGI, exploring how scalable intelligence, synthetic reasoning, and alignment strategies are paving a new path forward. This isn’t your average tech commentary—it’s a bold vision for conscious AI systems that reason, align, and adapt beyond narrow tasks.

🔍 Read it, upvote it if it sparks something, and let’s ignite a collective conversation about the future of AGI.

https://huggingface.co/blog/Parveshiiii/glimpses-of-agi

Parveshiiii

posted an update 5 months ago

Post

2855

🧠 MathX-5M by XenArcAI — Scalable Math Reasoning for Smarter LLMs

Introducing MathX-5M, a high-quality, instruction-tuned dataset built to supercharge mathematical reasoning in large language models. With 5 million rigorously filtered examples, it spans everything from basic arithmetic to advanced calculus—curated from public sources and enhanced with synthetic data.

🔍 Key Highlights:
- Step-by-step reasoning with verified answers
- Covers algebra, geometry, calculus, logic, and more
- RL-validated correctness and multi-stage filtering
- Ideal for fine-tuning, benchmarking, and educational AI

📂 - XenArcAI/MathX-5M

1 reply

AI & ML interests

Recent Activity

Team members 4

XenArcAI's activity

README

[bot] Conversion to Parquet