PulseGateWireMethodologyCompanyThe global software index— through the gate this hour
Coverage—in the index
Freshness—newest listing
Cadence—last week average · — today
Index9 markets91 categories · 140 niches
PulseGate

The global index of software taking shape now.

Stores show what passed through a store. Launch sites show what launched there. Catalogs show what entered their catalog. Each sees the market through its own gate. PulseGate reads across them.

FollowGitHubX (Twitter)LinkedIn
Platform
IndexIndex by setCategoriesWireMethodologySupply IndexDead ProjectsData SourcesCoverage RulesGlossaryEmbed Widget
Support
Help CenterSubmit your projectReport an Issue
Company
AboutTeamDimaxiaPress & DataContactPlatform Status
Legal
PrivacyTermsDisclaimer
Dimaxia · Dymaxio s.r.o. · Prague, Czechia · © 2026WatchlistSitemapSystem status
PulseGateDADataFlint
Visit↗
Skip to content
  1. Index›
  2. Autonomous agents & workflows›
  3. DataFlint
← Back to the index
DA

DataFlint

dataflint.io·Autonomous agents & workflows

DataFlint is an AI-powered platform for Apache Spark that enriches Spark logs and provides agentic tools to optimize job performance and reduce infrastructure costs. It includes multiple AI agents for job analysis, cluster management, and observability, targeting data engineers and enterprises running Spark workloads.

Commercial OSSApache-2.0WebCloud-managedAPI
DDataFlint preview
Visit dataflint.io↗
472stars
54forks
6features
2023since

Overview

6 features

DataFlint sits in PulseGate's Autonomous agents & workflows category. It focuses on reducing the complexity and cost of managing and optimizing Apache Spark jobs using AI agents. It is built as a B2B product for data engineers and enterprises using Apache Spark. It follows a commercial open-source model under the Apache-2.0 license. It runs on the web and API.

DataFlint builds and maintains DataFlint, and it first shipped in 2023. Development happens publicly on GitHub with 472 stars and 43 commits in the last 90 days. Key capabilities include spark log enrichment, AI agents, and job optimization. It exposes integrations via an MCP server.

Summary written by a language model from the project’s public pages.

  • ✓Spark log enrichment
  • ✓AI agents
  • ✓Job optimization
  • ✓Cost reduction
  • ✓Fleet observability
  • ✓Cluster management
Tags
spark-optimizationai-agentslog-analysis
AI capabilities
CodeStructured
Inference: Cloud API

Built with & integrations

Framework
Next.js
Hosting
Vercel
AI providers
multiple
Connectors
MCP
Runs on
BrowserCloud-managedAPI-only
Detected from
Next.js
x-nextjs-prerender header · /_next/static/ in the HTML · __next_f in the HTML
Vercel
x-vercel-id header · x-vercel-cache header

Trust & compliance

License
Apache-2.0
Verified signals
✓HTTPS✓Privacy Policy✓Terms of Service✓Open Source✓GitHub · ★ 472✓Active maintenance
Legal
Privacy Policy →Terms of Service →

Indexing history

1

What PulseGate has recorded for this listing

  1. Indexed25 Jun · 18:08 UTC
    Listing verified against its public source
    Source: PulseGate · Open ↗

Frequently asked questions about DataFlint

What is DataFlint?
DataFlint focuses on reducing the complexity and cost of managing and optimizing Apache Spark jobs using AI agents. It is catalogued under Autonomous agents & workflows on PulseGate.
Who is DataFlint for?
DataFlint is a B2B product built for data engineers and enterprises using Apache Spark.
Is DataFlint free?
The core is open source (Apache-2.0), with commercial offerings on top.
What platforms does DataFlint run on?
DataFlint runs on the web and API.
Is DataFlint still active?
The GitHub repository shows 43 commits in the last 90 days.
Who develops DataFlint?
DataFlint is developed by DataFlint.
When did DataFlint launch?
DataFlint first shipped in 2023.
Is DataFlint open source?
Yes — DataFlint is open source under the Apache-2.0 license, developed on GitHub.

At a glance

Pricing
Commercial OSS · pricing page detected
Platforms
Web
Languages
English
Open source
Yes · ★ 472
License
Apache-2.0
Built for
Data engineers and enterprises using Apache Spark
Model
B2B
Solves
Reducing the complexity and cost of managing and optimizing Apache Spark jobs using AI agents.

Registered as

GitHub
dataflint/spark

Developer

Team
↗ GitHub

Open source

View on GitHub →
Stars
472
Forks
54
Open issues
5
Last commit
2 Jun 2026
Commits 90d
43
Contributors
11
Authorship
Team
Default branch
main
Latest release
v0.9.9 · 18 May 2026

Index record

Identity confidence
High · 92.6
Indexed
25 Jun 2026
Lifecycle
Alive
Last seen
25 Jun 2026
Identity audit (12)
Slug
dataflint-production-aware-ai-agents-for-apache-spark-dataflint-io
Lifecycle last checked
8 Aug 2026
Verification state
Indexed for public listing
Listing state
Listed: yes
Index status
Included in index
Latest evidence snapshot
25 Jun 2026
Timeline basis
Indexed-at chronology. This listing's first-seen date was written by the catalog backfill, not observed here, so it is not treated as a sighting.
Name from
Derived from the project's own page and URL.
Category from
Recorded source: taxonomy-pass-a.
Summary from
Written by a language model from public pages.
Languages from
Detected by a language model, checked against the page's own declaration.
Canonical URL
https://dataflint.io/

Ship this? Send a correction — no account, and you get a link to follow it.

Similar projects

Closest matches by what these projects do

  • FAFlint AIflint.com
  • AGagentfluentgithub.com