Microsoft Access Batch

Microsoft Access Batch

Certified

Bulk insert rows into a Microsoft Access database using prepared statements

Reads ION-formatted data from Kestra internal storage and performs high-performance batch inserts into a Microsoft Access database via the UCanAccess JDBC driver. Data is processed in chunks (default 1,000 rows) to optimize memory and performance.

yaml
type: io.kestra.plugin.jdbc.access.Batch

Fetch rows from a query and bulk insert them into an Access table.

yaml
id: access_batch
namespace: company.team

tasks:
  - id: query
    type: io.kestra.plugin.jdbc.access.Query
    url: jdbc:ucanaccess:///myfile.accdb
    outputDbFile: true
    sql: |
      SELECT product_id, product_name, price
      FROM products
      LIMIT 1500
    fetchType: STORE

  - id: insert
    type: io.kestra.plugin.jdbc.access.Batch
    from: "{{ outputs.query.uri }}"
    accessFile: "{{ outputs.query.databaseUri }}"
    sql: INSERT INTO products_copy VALUES (?, ?, ?)
Properties

Input file from internal storage

URI of the source file (kestra://) containing rows to insert

The JDBC URL to connect to the database

Access database file (optional)

Optional URI to an existing Access database file stored in Kestra internal storage.

When provided, the file is downloaded into the task working directory and used as the Access database for the batch insert.

Default1000

Batch size per executeBatch call

Number of rows sent per JDBC batch before commit; default 1,000

SubTypestring

Columns bound to placeholders

Ordered column names matching ? placeholders; if omitted, placeholder count must match all columns in the input row

Default10

Maximum number of pooled connections

Maximum connections held in the pool for a given URL and credentials. Default 10. Increase for flows that run many concurrent queries against the same database to avoid waiting for an available connection. Ignored when connectionPooling is false or for embedded drivers.

Defaulttrue

Reuse database connections via a connection pool

When true (default), connections are pooled and reused across executions, keyed by URL and credentials, removing the connect and TLS-handshake cost on each run. Set to false if your SQL relies on session state persisting on the connection (for example SET search_path, session-scoped temp tables or variables), since pooled connections are reused. Embedded drivers (DuckDB, SQLite, MS Access) never pool regardless of this setting.

DefaultAUTO
Possible Values
AUTOSTREAMLOCAL

Input handling strategy

Controls how input is read during processing and retries. AUTO buffers small files locally (<= localBufferMaxBytes) and streams large files. STREAM always streams from internal storage. LOCAL always buffers input to a local temporary file before processing.

Default104857600

Maximum number of bytes buffered locally

Used by AUTO and LOCAL input handling. In AUTO, files larger than this threshold are streamed. In LOCAL, files larger than this threshold fail fast.

Default3

Maximum number of retries for transient failures

Retries are attempted only for transient failures such as temporary I/O and recoverable SQL errors.

DefaultV2003
Possible Values
V2000V2003V2007V2010V2016

Access database version to use when creating a new database file

Specifies the Microsoft Access file format used when UCanAccess creates a new database. Only applies when the target database file does not already exist. See https://spannm.github.io/ucanaccess/20-getting-started.html for details.

The database user's password

Reference (ref) of the pluginDefaults to apply to this task.

Defaulttrue

Resume from the last successfully committed chunk on retry

DefaultPT1S

Delay between retry attempts

Uses ISO-8601 duration format, for example PT1S.

DefaultINPUT
Possible Values
NONEINPUTALL

Controls which failures are retried

INPUT retries input handling failures, ALL retries all retryable failures.

Parameterized INSERT statement to execute

Prepared INSERT with ? placeholders for each bound column. Example: INSERT INTO VALUES (?, ?, ?) for three columns; use column list if inserting a subset

Table used to auto-discover columns

Retrieves column names from the given table when columns is empty. If sql is also omitted, an INSERT statement is generated automatically using the discovered columns

The time zone id to use for date/time manipulation. Default value is the worker's default time zone id

The database user

Total rows read

Rows inserted or updated

Unitqueries

The number of batch queries executed.

Unitrecords

The number of records processed.

Unitrecords

The number of records updated.