Home / Companies / Fivetran / Blog / March 2024

March 2024 Summaries

13 posts from Fivetran

Filter
Month: Year:
Post Summaries Back to Blog
Fivetran, a data integration platform, shifted its sales-led strategy to a product-led approach to reduce costs and focus on high-complexity customers. The company A/B tested a new operating motion, which involved creating a cross-functional growth team with a growth mindset, establishing processes and rituals, using data to drive decisions, and removing sales from certain accounts. The experiment showed that the product-led cohort had higher conversion rates, but lower average revenue per account (ARPA). Despite this, the company still saw significant efficiency gains, which led to an evidence-based business decision to expand the product-led motion more broadly in its go-to-market strategy.
Mar 29, 2024 900 words in the original blog post.
The European Union has passed the EU AI Act, a comprehensive piece of legislation aimed at regulating artificial intelligence systems within the EU. The regulation seeks to establish a harmonized framework for the development, deployment, and use of AI technologies across various industry verticals while ensuring compliance with fundamental rights and values. The act applies to AI systems placed in the EU market or used within the EU, categorizes them based on risk level, and requires high-risk AI systems to adhere to strict requirements. Companies will need to ensure data quality, transparency, and privacy, develop or enhance their data governance frameworks, and conduct risk assessments for their AI systems. The act also establishes an oversight mechanism and enforcement measures to ensure compliance, including penalties for non-compliance. Fivetran, a leading data integration platform, can help organizations align with the EU AI Act by providing scalable control solutions, top-tier security, and complete visibility into the entire data movement process.
Mar 29, 2024 1,185 words in the original blog post.
Powell Industries is revolutionizing industrial electrical systems by leveraging data, teamwork, and advanced analytics to achieve better processes, higher product quality, and strategic innovation. Ajay Bidani, Data and Insights Manager at Powell, emphasizes the importance of data democratization, enabling data analysis and decision-making across all departments, and developing critical soft skills in data engineering. He highlights the need for a strong and reliable data foundation to support AI and machine learning success, as most problems with AI start with data issues such as data quality and availability. Bidani's approach focuses on building an inclusive data culture, training employees to make data-informed decisions, and encouraging open-mindedness and listening in modern data professionals.
Mar 28, 2024 498 words in the original blog post.
An automated helpdesk chatbot can be created using a combination of text-rich data sources, data integration solutions, and scripting languages. This process is exemplified by leveraging Zendesk data hosted in an S3 data lake, using Python, OpenAI’s GPT-4, ChromaDB for a vector database, and the LangChain API. The workflow begins with setting up a Fivetran connector to sync Zendesk data to a data lake or warehouse. The next steps involve extracting, transforming, and vectorizing data to load it into the vector database, followed by setting up a user interface for prompt acceptance. The retrieval model is built by vectorizing and storing data in a database, allowing for RAG operations that search the database for relevant information and augment user prompts before processing with the LLM. Streamlit is used to create a front end for user interaction, and the workflow is designed for easy updates and comprehensive responses as the vector database is enriched. The potential for further innovation includes streamlining the process, integrating functionalities, and using knowledge graphs, highlighting the evolving landscape of generative AI and data integration technologies.
Mar 26, 2024 1,253 words in the original blog post.
Fivetran continues to expand its cloud regions library for global businesses' data processing needs while meeting data residency requirements. The latest addition is Azure Japan, bringing the total number of Azure data centers supported by Fivetran to ten worldwide. Additionally, Fivetran offers a Business Critical plan for Azure users with enhanced security features and access to HVR, a high-volume database replication solution. Companies can leverage Fivetran's multi-cloud functionality for more flexibility and control over their data movement across GCP, AWS, and Azure cloud regions worldwide.
Mar 21, 2024 255 words in the original blog post.
Fivetran Cloud Function Connectors` are a powerful solution for large international enterprises to streamline data integration, reduce costs and boost productivity by leveraging `custom code orchestrated and maintained by Fivetran`, utilizing serverless platforms like AWS Lambda or Azure Functions. They can be used when there isn't a native Fivetran connector for your data source, providing more flexibility and faster implementation compared to custom connectors. These connectors are valuable for supporting sources that Fivetran does not natively support, as well as use cases such as enhancing existing connectors with additional data, in-flight transformations of data, complex pipelines, external orchestration, filtering to extract only a subset of data from a given source, and optimizing data pipelines through source-side filtering. This solution helps reduce costs and boost productivity by streamlining data integration and providing the necessary tools for data science, sharing and monetization projects throughout the Modern Data Stack environment.
Mar 20, 2024 585 words in the original blog post.
The study highlights the importance of high-quality data in AI success, as poor data quality can lead to misinformed business decisions that impact an organization's global annual revenue by 6%, or $406 million on average. Despite optimism about AI technologies, many organizations struggle to utilize AI and human intervention is still widely used due to siloed, low-quality, and stale data. To overcome these challenges, companies need to invest in tools that strengthen data movement, governance, and security, and adopt a strong data foundation through data integration.
Mar 20, 2024 774 words in the original blog post.
Fivetran has honored its top partners for their exceptional achievements in providing data solutions, empowering businesses worldwide to unlock the potential of data analytics. The awards celebrate the dedication and expertise that these partners bring to the table, helping companies transform data into actionable insights for strategic decision-making. Among the winners are Databricks as Technology Partner of the Year, AWS as Innovation Technology Partner, Snowflake as Strategic Alliance Partner of the Year, dbt Labs as Better Together Technology Partner, Sigma Computing as Ecosystem Partner of the Year, and several GSI and SI partners for their outstanding performance in different regions.
Mar 19, 2024 1,763 words in the original blog post.
There's unprecedented industry buzz around generative AI, but for companies trying to deploy it, there are several key challenges. Most companies are still working on building the necessary blocks, and many are struggling with data quality and governance. The hardest part of AI is often the data, and companies need to carefully consider whether their use case would be better served by simpler predictive modeling or heuristics. To overcome these challenges, companies like Databricks and Fivetran are offering solutions such as retrieval augmented generation (RAG) that can help ensure factual and relevant outputs from generative AI models. By addressing data ownership, governance, and quality control, companies can unlock the full potential of generative AI and bring business value to their operations.
Mar 14, 2024 785 words in the original blog post.
Fivetran, Databricks, and AutoML streamline the process of building machine learning applications in the Databricks Lakehouse by automating data movement and efficient model creation. The text explains how to set up a relational database connector to the Databricks Lakehouse using Fivetran and move a wine quality dataset over for classification experiments and predicting wine quality based on various parameters. It also highlights the importance of having high-quality, usable data and how Fivetran's automated data platform helps achieve this by centralizing data and modernizing data infrastructure. The text concludes with an overview of managing source changes and schema drift, starting the initial sync from PostgreSQL to the Databricks Lakehouse, and building a wine quality application using Databricks AutoML.
Mar 12, 2024 2,413 words in the original blog post.
The text discusses the importance of prompt design and engineering in using foundation models for productivity or supplementing them with proprietary data. Good prompts should minimize ambiguity, provide context, and follow principles such as giving directions, providing examples, formatting responses, dividing labor, and evaluating outputs iteratively. Poor prompt design can lead to misleading results, intellectual property theft, and malicious applications like hallucinations and prompt injections. To minimize hallucinations, fact-check results, ensure high-quality training sets, and use knowledge graphs for semantic relationships. Protecting generative AI involves data governance, security, anonymizing sensitive data, using access control policies, sanitizing prompts, screening outputs, and allowing end users to report bad results.
Mar 06, 2024 1,215 words in the original blog post.
This article emphasizes the importance of data governance and security in preventing misuse of data, particularly with the increasing use of generative AI and other data products and technologies. Ensuring proper data governance and security relies on capabilities offered by software and platforms constituting an organization's data infrastructure, which are essential for tracking data, limiting access to necessary stakeholders, and maintaining scalability as the organization grows. The principles of data governance include observability, control, and scalability, while those of data security involve preventing unauthorized access to sensitive data through practices such as end-to-end encryption and purging data once it is no longer needed. A secure data infrastructure ensures the safety and integrity of data that feeds AI initiatives, enabling organizations to trust their models and bring them to market with confidence.
Mar 05, 2024 866 words in the original blog post.
A report by technology research firm GigaOm has compared the replication latency and total cost of ownership between Fivetran HVR and Qlik Replicate, with results favoring Fivetran HVR. The tests were performed on both platforms to compare latency and total cost of ownership across multiple change data volumes when replicating from an Oracle database source to a Snowflake cloud data warehouse destination. Fivetran HVR showed 27 times lower latency than Qlik Replicate at 200GB/hour of change data, and is 25% less expensive than Qlik Replicate at the same rate. The report highlights Fivetran's commitment to improving its database replication capabilities and addressing the needs of large enterprises utilizing their data for business insights, predictive analytics, and AI/ML workloads.
Mar 04, 2024 747 words in the original blog post.