Skip to main content
FivetranCustomer feedback

In a discussion about building a custom EL tool, a commenter suggests there is a market gap because incumbent tools like Fivetran often handle schema drift poorly and struggle with recovery and...

What happened

In a discussion about building a custom EL tool, a commenter suggests there is a market gap because incumbent tools like Fivetran often handle schema drift poorly and struggle with recovery and idempotency, leading to duplicate rows after failed loads.

Source

Post

Hi Community, I'll try not to sound like a sales person here. My team and I built this tool for us to use to move data from major databases(mysql, postgres, sql server mainly) to snowflake(I know it's pretty limited, but that's enough for us for now). Just want to know if the community thinks it's a good idea to open it up to the community for free or it's just unique cases just for us. It's mainly to solve some of the pains you might have encountered as well when doing EL to snowflake as a data engineer or architect. Anyway, things that bothers me and my team a lot were as follows, you can comment on how you solve this instead of writing a tool to do the syncing(and make us look stupid of course) Big database table, 1+ billion records syncing - if you've done this scale before, you know what I'm talking about, it's super slow. Export existing views instead of tables!!! - I don't see why the existing tools would not do that. Schedule per table! - Why do I have to run the syncing to the WHOLE database at scheduled time.... it doesn't make sense. Delete detections - Seriously, why this is so hard to achieve... SOX Audit - it's a perfect use case for the EL tool IMO, why isn't there one already. etc... These are some main pain points we faced during our operations, wanna see your thoughts as to how you solve these problems, and if our in house tool that solved these...

Keep reading with a free account

The rest of this post, and every signal for Fivetran, is in your free account.

Extracted from these lines

  • [comment u/Muted_Jellyfish_6784] There is room even with Fivetran and Airbyte out there, the gap is usually schema drift handling and how cleanly the tool recovers from a source changing shape mid-sync without dropping columns.

  • [comment u/Muted_Jellyfish_6784] Stress test idempotency, what happens on a replay after a failed load, duplicate rows from a bad retry are the top complaint about EL tools in production.

Comments on the post

5 of 11 comments
  • “Before you wrote your own ETL tool, which of the many existing ETL tools did you consider and why did you reject them? What does your tool provide that none of the existing tools provides?”

    u/NW19696 points · Sep 21, 2026View

  • “CDC brother CDC replication”

    u/Defiant_Month_4973 points · Sep 21, 2026View

  • “There is room even with Fivetran and Airbyte out there, the gap is usually schema drift handling and how cleanly the tool recovers from a source changing shape mid-sync without dropping columns. Stress test idempotency, what happens on a replay after a failed load, duplicate rows from a bad retry are the top complaint about EL tools in production. Document how it handles late-arriving data too, th”

    u/Muted_Jellyfish_67843 points · Sep 21, 2026View

  • “The most useful next step would be a reproducible benchmark against a staged Parquet load plus CDC, using the same dataset. Initial-load speed matters, but recovery after interruption, snapshot-to-CDC consistency, delete handling, reconciliation, and credential security are probably the real differentiators. Those results would make it much easier to tell whether this fills a genuine gap or mainly”

    u/PrimeWilliam1 points · Sep 21, 2026View

  • “Very cool! Are there any major advantages over using something like Openflow?”

    u/RazzmatazzReal55961 points · Sep 21, 2026View

Extracted by Autobound

From the Signal API record
Signal
Customer feedback

What this signalsUser posts often show product pain before it reaches reviews or churn.

Subreddit
r/snowflake

Companies

  • AirbyteAlso named
  • SnowflakeAlso named

The full record

From the Signal API record

Numbers

Mentions
2

Details

Timing
Ongoing state
Category
Reliability
Virality
Medium
Post kind
Text
Prominence
Aside
Company's role
Vendor

Topics and mentions

Topics

  • data integration
  • reliability
  • schema drift
  • elt

Extraction

Sentiment
Negative
Detected
Sep 21, 2026
signal_type
reddit-company
signal_subtype
customerFeedback

Use this data

Get every Reddit signal for Fivetran and the companies you sell to, in the tools you already use.

  1. Ask Claude about it

    Connect Autobound to Claude, Claude Code or Cursor with MCP. Then ask: “What changed at Fivetran this week?”

  2. Send it to your own tools

    The Signal API returns Reddit signals for any list of companies as JSON, for your CRM, warehouse or app.

  3. Try it free

    Sign up and spend your free credits on the companies you sell to.

    Start Free1,000 free credits

The API returns more than this page shows

This page shows a preview. The full reddit-company record in the Signal API and MCP can also have these 8 fields. Some fields are empty for some signals.

Company

  • linkedin_urlValue in the API
  • industriesValue in the API
  • employee_count_lowValue in the API
  • employee_count_highValue in the API
  • revenueValue in the API
  • descriptionValue in the API

Signal

  • signal_nameValue in the API
  • associationValue in the API
Show the full JSONThe record on this page and the API request

GET /v1/signals/cfb266cb-887a-5593-abd4-4ef1f7ccd50c returns this record as JSON. POST /v1/companies/enrich returns every signal for fivetran.com.

{
  "signal_id": "cfb266cb-887a-5593-abd4-4ef1f7ccd50c",
  "signal_type": "reddit-company",
  "signal_subtype": "customerFeedback",
  "detected_at": "2026-09-21T03:12:46+00:00",
  "company": {
    "name": "Fivetran",
    "domain": "fivetran.com"
  },
  "data": {
    "nsfw": false,
    "stage": "none",
    "awards": 0,
    "timing": "ongoing_state",
    "topics": [
      "data integration",
      "elt",
      "reliability",
      "schema drift"
    ],
    "post_id": "1wm16cg",
    "summary": "In a discussion about building a custom EL tool, a commenter suggests there is a market gap because incumbent tools like Fivetran often handle schema drift poorly and struggle with recovery and idempotency, leading to duplicate rows after failed loads.",
    "category": "reliability",
    "comments": [
      {
        "url": "https://www.reddit.com/r/snowflake/comments/1wm16cg/comment/pb4p5qg/",
        "depth": 0,
        "score": 6,
        "author": "NW1969",
        "excerpt": "Before you wrote your own ETL tool, which of the many existing ETL tools did you consider and why did you reject them? What does your tool provide that none of the existing tools provides?",
        "posted_at": "2026-09-21T09:25:30.000Z",
        "author_url": "https://www.reddit.com/user/NW1969/"
      },
      {
        "url": "https://www.reddit.com/r/snowflake/comments/1wm16cg/comment/pb3dru0/",
        "depth": 0,
        "score": 3,
        "author": "Defiant_Month_497",
        "excerpt": "CDC brother CDC replication",
        "posted_at": "2026-09-21T03:18:57.000Z",
        "author_url": "https://www.reddit.com/user/Defiant_Month_497/"
      },
      {
        "url": "https://www.reddit.com/r/snowflake/comments/1wm16cg/comment/pb6l03b/",
        "depth": 0,
        "score": 3,
        "author": "Muted_Jellyfish_6784",
        "excerpt": "There is room even with Fivetran and Airbyte out there, the gap is usually schema drift handling and how cleanly the tool recovers from a source changing shape mid-sync without dropping columns. Stress test idempotency, what happens on a replay after a failed load, duplicate rows from a bad retry are the top complaint about EL tools in production. Document how it handles late-arriving data too, th",
        "posted_at": "2026-09-21T15:40:18.000Z",
        "author_url": "https://www.reddit.com/user/Muted_Jellyfish_6784/"
      },
      {
        "url": "https://www.reddit.com/r/snowflake/comments/1wm16cg/comment/pb4m0d1/",
        "depth": 0,
        "score": 1,
        "author": "PrimeWilliam",
        "excerpt": "The most useful next step would be a reproducible benchmark against a staged Parquet load plus CDC, using the same dataset. Initial-load speed matters, but recovery after interruption, snapshot-to-CDC consistency, delete handling, reconciliation, and credential security are probably the real differentiators. Those results would make it much easier to tell whether this fills a genuine gap or mainly",
        "posted_at": "2026-09-21T08:58:16.000Z",
        "author_url": "https://www.reddit.com/user/PrimeWilliam/"
      },
      {
        "url": "https://www.reddit.com/r/snowflake/comments/1wm16cg/comment/pb6cd3u/",
        "depth": 0,
        "score": 1,
        "author": "RazzmatazzReal5596",
        "excerpt": "Very cool! Are there any major advantages over using something like Openflow?",
        "posted_at": "2026-09-21T15:03:47.000Z",
        "author_url": "https://www.reddit.com/user/RazzmatazzReal5596/"
      }
    ],
    "evidence": [
      "[comment u/Muted_Jellyfish_6784] There is room even with Fivetran and Airbyte out there, the gap is usually schema drift handling and how cleanly the tool recovers from a source changing shape mid-sync without dropping columns.",
      "[comment u/Muted_Jellyfish_6784] Stress test idempotency, what happens on a replay after a failed load, duplicate rows from a bad retry are the top complaint about EL tools in production."
    ],
    "virality": "medium",
    "post_date": "2026-09-21T03:12:46.000Z",
    "post_kind": "text",
    "post_text": "Hi Community, I'll try not to sound like a sales person here. My team and I built this tool for us to use to move data from major databases(mysql, postgres, sql server mainly) to snowflake(I know it's pretty limited, but that's enough for us for now). Just want to know if the community thinks it's a good idea to open it up to the community for free or it's just unique cases just for us. It's mainly to solve some of the pains you might have encountered as well when doing EL to snowflake as a data engineer or architect.\n\nAnyway, things that bothers me and my team a lot were as follows, you can comment on how you solve this instead of writing a tool to do the syncing(and make us look stupid of course)\n\nBig database table, 1+ billion records syncing - if you've done this scale before, you know what I'm talking about, it's super slow.\n\nExport existing views instead of tables!!! - I don't see why the existing tools would not do that.\n\nSchedule per table! - Why do I have to run the syncing to the WHOLE database at scheduled time.... it doesn't make sense.\n\nDelete detections - Seriously, why this is so hard to achieve...\n\nSOX Audit - it's a perfect use case for the EL tool IMO, why isn't there one already.\n\netc...\n\nThese are some main pain points we faced during our operations, wanna see your thoughts as to how you solve these problems, and if our in house tool that solved these...",
    "sentiment": "negative",
    "subreddit": "snowflake",
    "post_title": "My team and I built a EL tool especially for snowflake, not sure if we are reinventing the wheels or it's worth sharing with the community. let me know what you think.",
    "prominence": "aside",
    "source_url": "https://www.reddit.com/r/snowflake/comments/1wm16cg/my_team_and_i_built_a_el_tool_especially_for/",
    "entity_role": "vendor",
    "post_author": "loantochoose",
    "upvote_ratio": 0.9,
    "mention_count": 2,
    "mention_surge": false,
    "subreddit_url": "https://www.reddit.com/r/snowflake/",
    "total_upvotes": 8,
    "comments_total": 11,
    "total_comments": 11,
    "other_companies": [
      {
        "name": "Airbyte",
        "role": "competitor",
        "domain": "airbyte.com"
      },
      {
        "name": "Snowflake",
        "role": "partner",
        "domain": "snowflake.com"
      }
    ],
    "post_author_url": "https://www.reddit.com/user/loantochoose/",
    "signal_category": "feedback",
    "comments_included": 5
  }
}

Long text fields are shortened on this page.

Looking up one signal by its id is free. Enrich costs 2 credits per signal returned; a call with no results is free.