Skip to content

Follow-up to #341: next readers set from DB-Engines ranking #570

Description

@thomasp85

SedonaDB is pretty high priority as it will provide a pure rust, in-memory backend with spatial support built in. An ADBC driver for it is in the works.

Cross-referencing the current 17 dialects against the DB-Engines ranking (438 systems, Sept 2026) shows the remaining gaps fall into two cost classes. All relational systems in the top 10 are already covered (Oracle, MySQL, MSSQL, PostgreSQL, Snowflake, Databricks, SQLite).

Level 1 — Wire-protocol aliases (no new dialect module)

These systems speak an existing wire protocol, so support is a scheme mapping in detect_dialect/dialect_for_scheme plus golden tests — no new dialect module, no new driver machinery. Several outrank systems we already support.

Rank System Score Maps to Notes
15 Microsoft Azure SQL Database 70.8 mssql Ranked above BigQuery (#19) and ClickHouse (#26), both supported
59 Presto 6.2 trino Same engine family
32 / 36 Microsoft Fabric / Azure Synapse 19.4 / 16.2 mssql family T-SQL warehouse endpoints; verify dialect deltas
28 Apache Spark SQL 23.3 databricks-ish Spark Thrift Server; check overlap with databricks dialect
44 / 63 / 64 / 72 / 116 Aurora, TimescaleDB, Greenplum, CockroachDB, YugabyteDB 9.6 / 5.9 / 5.6 / 4.2 / 2.2 postgres Thin subclasses may be needed: CockroachDB temp-table limits, Greenplum's older PG base
82 / 89 / 146 TiDB, SingleStore, StarRocks 3.3 / 3.1 / 1.6 mysql Also PlanetScale, Vitess, Apache Doris
85 QuestDB 3.1 postgres (pgwire)

Suggested acceptance criteria: scheme aliases resolve to the right dialect, golden SQL tests pass for each alias, and each gets at least a smoke-test row in the Tier 2/3 live matrix where a container exists.

Level 2 — New dialect modules

Ranked by popularity. None have ADBC drivers, but all ship vendor ODBC drivers, so each rides the existing ODBC fallback: dialect module + golden tests + live config, no new driver infrastructure.

Rank System Score Dialect notes
9 IBM Db2 110.6 Largest relational gap by ~40 points; FETCH FIRST n ROWS ONLY limit syntax, ANSI double-quote idents
16 Apache Hive 70.1 HiveQL; backtick quoting; large legacy installed base
22 SAP HANA 33.2 Enterprise analytics
23 Teradata 31.2 Distinct quirks (SAMPLE, TOP variant)
25 SAP Adaptive Server (Sybase) 26.1 T-SQL ancestor; may partially reuse mssql dialect
41 Apache Impala 11.1 Hive-family
51 Vertica 7.7 Analytic workloads, mostly ANSI-ish
39 / 60 Informix / Netezza 12.9 / 6.0 Legacy enterprise, declining — lowest priority

Db2 is the obvious first pick. A reasonable sequencing: Db2 alone first (validates the ODBC-fallback dialect workflow end-to-end), then the Hive/Impala pair (shared dialect groundwork), then HANA/Teradata/Vertica as demand warrants.

Activity

  1. ianmcook commented on Oct 2, 2026

    @ianmcook

    We have ADBC drivers available for most of these:

    System Compatible ADBC driver Notes
    Microsoft Azure SQL Database https://adbc-drivers.org/drivers/mssql/
    Presto https://adbc-drivers.org/drivers/presto/
    Microsoft Fabric / Azure Synapse https://adbc-drivers.org/drivers/mssql/
    Apache Spark SQL https://adbc-drivers.org/drivers/spark/
    Aurora, TimescaleDB, Greenplum, CockroachDB, YugabyteDB https://arrow.apache.org/adbc/current/driver/postgresql.html Tested with TimescaleDB, CockroachDB, YugabyteDB; see https://github.com/columnar-tech/adbc-quickstarts/tree/by-database and [1]
    TiDB, SingleStore, StarRocks https://adbc-drivers.org/drivers/mysql/ Tested with TiDB and Vitess; see https://github.com/columnar-tech/adbc-quickstarts and [2]
    QuestDB coming soon Ask Ian for details
    IBM Db2 IBM has been nonresponsive recently
    Apache Hive Is there actually demand for this? If so we might have a solution
    SAP HANA https://docs.columnar.tech/drivers/sap-hana/ We're prepared to make this free and open source if someone at SAP could please to talk to us
    Teradata https://docs.columnar.tech/drivers/teradata/ We're prepared to make this free and open source if someone at Teradata could please to talk to us
    SAP Adaptive Server (Sybase)
    Apache Impala Is there actually demand for this? If so we might have a solution
    Vertica coming soon Ask Ian for details
    Informix / Netezza

    [1] For SingleStore, use the SingleStore ADBC driver: https://adbc-drivers.org/drivers/singlestore/
    [2] For Doris, use the Flight SQL ADBC driver: https://arrow.apache.org/adbc/current/driver/flight_sql.html

  2. thomasp85 commented on Oct 9, 2026

    @thomasp85
    CollaboratorAuthor

    Also chDB to complete the ClickHouse story

  3. thomasp85 commented on Oct 9, 2026

    @thomasp85
    CollaboratorAuthor

    @ianmcook as for the question of demand, the first round just merged in included all the backends that we have received so far so everything on top of that is just icing for now. That being said we uncovered so many portability issues in our current model through the expansion so it can be good to expand into some weird dialects just for the sake of it

  4. ianmcook commented on Oct 9, 2026

    @ianmcook

    Also chDB to complete the ClickHouse story

    chDB is available as an ADBC driver: https://clickhouse.com/docs/chdb/install/adbc

  5. thomasp85 commented on Oct 9, 2026

    @thomasp85
    CollaboratorAuthor

    Yup - that was the plan. It'll join DataFusion as an ADBC-backed, in-process engine

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions