npx skills add ...
npx skills add celigo/ai --skill configuring-imports
Configure Celigo imports -- the destination step that writes records to external systems. Use when creating imports, choosing the adaptor type, setting up field mappings, lookups, upsert logic, AI agent imports, or file-based imports.
npx skills add celigo/ai --skill configuring-imports
An import is the data destination in a Celigo integration. It takes records from an upstream step and writes them to an external system -- REST APIs, databases, ERPs, file servers, or AI models. Every import is bound to exactly one connection and one adaptor type.
Imports handle six concerns:
mappings[] array) by default; NetSuite and Salesforce imports only support Mapper 1.0 (mapping.fields[] / mapping.lists[])oneToMany: true and pathToMany to the child array path (e.g., "lineItems") when one source record should create multiple import operationspageProcessors[] entry, but planned when building the import. The response is available via _json (the raw API response) and errors. Use _json.fieldName to extract from the response (e.g., _json.id for a created record's ID, _json.output.1.content.0.text for OpenAI responses). Response mapping uses Transformation 1.0 syntax (extract/generate pairs), not the newer expression-based transformspageProcessors[] entry, but planned when building the import. Use to transform or enrich the merged record before downstream stepsImports are used across flows, APIs, and tools.
When records arrive at an import step, this pipeline executes in strict order:
pageProcessors[] entry)pageProcessors[] entry, not on the import itselfKey distinction: Response mapping lives on the flow's pageProcessors[] entry, not on the import resource. When building an import that needs to pass data downstream, plan the response mapping at flow design time.
Submit structured records to APIs, databases, or ERPs. The vast majority of imports.
NetSuiteDistributedImport -- high-performance SuiteApp writes (add, update, addupdate, delete, attach, detach)HTTPImport -- REST/GraphQL APIs (POST, PUT, PATCH, DELETE). Supports connector-assisted (formType: "assistant") and GraphQL (graph_ql) modesSalesforceImport -- Salesforce CRUD via SOAP, REST, Bulk, or Composite Record APIRDBMSImport -- SQL databases (Snowflake, PostgreSQL, MySQL, SQL Server, Oracle). Uses per_record, bulk_insert, or bulk_load query typesMongodbImport, DynamodbImport, JDBCImport -- other databasesWrite or upload files to remote storage. Require the file{} configuration block. Two modes:
file{} block defines the output format.blobKeyPath field to the path in the record that contains the blob key.Adaptor types:
HTTPImport with http.type: "file" -- upload files over HTTP to cloud storage APIs (Google Drive, Box, Dropbox, Azure Blob Storage)FTPImport -- CSV, XML, JSON, XLSX, EDI files to FTP/SFTPS3Import -- objects to Amazon S3AS2Import -- AS2 EDI file transmissionFileSystemImport -- local/on-premise filesystem writesInvoke AI models for classification, extraction, or safety checks. No _connectionId required unless using BYOK.
AiAgentImport -- OpenAI or Gemini model invocations with structured output, tool use, and reasoningGuardrailImport -- PII detection, content moderation, or custom AI-based validationWrapperImport -- custom pre-built stack connectors (Walmart, BigCommerce)ToolImport -- invoke a Celigo Tool resourceThe core decision on most record-based imports is the operation -- what happens to each record at the destination. Create requires no match key; update and delete require one; upsert checks first and does whichever applies. Users describe the intent in business terms ("look up the customer and update them", "match by email and upsert", "skip the ones that already exist") that resolve into five behaviors:
When matching applies, decide four things: the matching behavior, the match key field(s) (email, external_id, customer.id -- required for anything other than always-create), the on-match action (update, skip, fail), and the on-no-match action (create, skip, fail).
How a destination implements matching is adaptor-specific -- there is no single "matching mode" field. A destination might expose: a native upsert keyed off an external ID (Salesforce upsert, NetSuite upsert, RDBMS ON CONFLICT); a distinct addupdate operation that handles both paths in one call; an ignoreExisting flag paired with a lookup that probes before writing; two separate create and update endpoints with no upsert variant (see composite imports below); or a lookup endpoint plus separate create and update endpoints, where the lookup runs pre-write to drive the create-vs-update decision. Some destinations have no matching concept at all -- writing a CSV to FTP, sending an email, posting to a webhook -- so every record goes out as-is.
Prefer import-level matching over a separate lookup step. Imports natively support this pre-write check, so a single import that does the matching and the write together means fewer steps, fewer round-trips, and no glue logic to maintain. A standalone lookup earns its place only when the looked-up data has a consumer beyond the write -- a router branching on something other than "does this exist", an AI agent reasoning over the result, or multiple downstream steps reading different fields. If the only consumer is the destination call itself, the work belongs inside the import.
When an HTTP destination has no native upsert but exposes separate create and update endpoints, a composite import pins both endpoints on a single import node, role-tagged create and update. The runtime picks per record via a match-key check -- the same way a native upsert would -- so the flow stays one step. Prefer this over two separate imports driven by an upstream lookup or router; reach for separate imports only when the create and update paths must diverge beyond endpoint selection (different mappings, different downstream consumers, or different hook chains).
| Your data goes to... | Use adaptorType | Category | Read schema |
|---|---|---|---|
| REST or GraphQL API | HTTPImport | Record-based | http.yml |
| NetSuite (any method) | NetSuiteDistributedImport | Record-based | netsuitedistributed.yml |
| Salesforce objects | SalesforceImport | Record-based | salesforce.yml |
| SQL database (Snowflake, PostgreSQL, etc.) | RDBMSImport | Record-based | rdbms.yml |
| MongoDB | MongodbImport | Record-based | mongodb.yml |
| DynamoDB | DynamodbImport | Record-based | dynamodb.yml |
| JDBC database (non-built-in) | JDBCImport | Record-based | jdbc.yml |
| Files over HTTP (Google Drive, Box, Dropbox, Azure Blob) | HTTPImport with http.type: "file" | File-based | http.yml |
| Files to FTP/SFTP | FTPImport | File-based | ftp.yml |
| Files to S3 | S3Import | File-based | s3.yml |
| AS2 EDI transmission | AS2Import | File-based | as2.yml |
| Local filesystem | FileSystemImport | File-based | filesystem.yml |
| OpenAI / Gemini | AiAgentImport | AI | aiagent.yml |
| PII detection / content moderation | GuardrailImport | AI | guardrail.yml |
| Celigo Tool | ToolImport | Tool | wrapper.yml |
| Pre-built stack connector | WrapperImport | Stack | wrapper.yml |
Raw HTTP is the fallback, not the default. Pick the most specific match, in order:
HTTPImport against that app's REST API.HTTPImport, but it runs on a connector-backed connection and takes its endpoint config from the connector.adaptorType is case-sensitive: NetSuiteDistributedImport, not netsuitedistributedimport.
Every import needs at minimum: name, adaptorType, _connectionId (except AiAgentImport/GuardrailImport without BYOK), and the adaptor config block (http{}, netsuite_da{}, salesforce{}, etc.).
All schemas are in references/schemas/:
What system are you writing data to? This determines adaptor type, connection type, and configuration shape.
Before building from scratch, look at what already exists:
The account index auto-refreshes when stale (>4 hours). Force a fresh snapshot with celigo account snapshot.
Always run this check before writing any HTTP config. Celigo maintains 550+ HTTP connectors with pre-configured auth, endpoints, and resources. Hand-write a manual HTTPImport from public API docs only when this search comes up empty or the connector doesn't cover the operation you need.
If a connector exists, create the connection from it (http._httpConnectorId -- see configuring-connections > Check for a pre-built connector and global iClient) and take the import's relativeURI, method, and body/response shapes from the connector's endpoint metadata rather than reconstructing them from public API docs. The connector-reference fields on the import itself (http._httpConnectorEndpointId, http._httpConnectorVersionId, http._httpConnectorResourceId) are read-only -- the platform sets them; what you control is the connection and the endpoint config you copy from the connector.
For NetSuite, Salesforce, and RDBMS connections, discover available record types and fields:
metadata fields returns field IDs, types, and groups — use the IDs for mapping.fields[].generate and mapping.lists[].fields[].generate. Sublist names (e.g., "item", "addressbook") appear as groups, which map to mapping.lists[].generate. Lookup field IDs here are the searchField/resultField values for netsuite_da.lookups[].metadata fields returns field API names, types, and relationship info. Use field API names for salesforce.sObjectType lookups and for discovering which fields are createable/updateable.metadata fields returns column names and types for a table — use these to write SQL queries (see writing-sql) and verify column names before building bulkInsert.tableName or bulkLoad.tableName.Is this a record-based import (submit records to an API/database/ERP), a file-based import (write files to storage), or an AI import (invoke a model)?
Refer to the Adaptor Decision Matrix in the Quick Reference above.
Reference the Schema Index for the exact fields needed. Use the Which Schemas to Read decision rule to determine which files to consult.
Some destination APIs accept files only as multipart/form-data POSTs (Jira attachments, QuickBooks attachables, OpenAI file uploads). This is an HTTPImport in file-transfer mode (http.type: "file") and works nothing like a JSON record write. Four pieces have to line up:
{"extract": "data.0.blobKey", "generate": "blobKey"}. The record then carries a blobKey (a storage pointer, not content).multipart/form-data; success/error response media types usually override to JSON so replies parse normally.{name, value, type} plus optional filename (include the extension) and mime-headers. The file part is "type": "attachment" with "value": "{{blob}}" -- {{blob}} is the only accepted value for an attachment (anything else 422s). Fields the API wants alongside the file ride as "type": "inline" parts; an inline part whose value is a JSON object must be serialized. One file reference per import.blobKeyPath (Advanced settings) -- the JSON path in the record where the blobKey lives (blobKey, or file.blobKey if nested). At send time the platform follows it into blob storage and streams the real bytes into the attachment part.The platform assembles the final MIME body itself -- it generates the boundary (never hardcode one), writes each part's Content-Disposition, and substitutes the attachment part with the raw file bytes. The parts array is a build recipe, not the payload.
Not every multipart API is form-data. Some upload endpoints expect multipart/related instead (e.g. Google Drive's /upload/drive/v3/files?uploadType=multipart), and the parts-array machinery does NOT apply. There the request body is the literal MIME document: explicit boundary, a JSON metadata part, and a content part referencing {{blob}} (double braces). Check which flavor the destination API documents before building -- mixing them produces an import that saves cleanly and fails at runtime.
Some destination APIs only acknowledge a write (an HTTP 202, a job ticket) and finish it in the background -- bulk loads, file ingestion, document conversion. Attach an async helper to the import (http._asyncHelperId) so the step submits, polls a status export until the external work completes, and only then resolves. The mechanics and constraints (a status export with done/error value lists and poll intervals; no transform, output filter, or hook on the async-configured step; helpers cannot nest) are identical to the export side -- see configuring-exports > Async APIs (submit, poll, fetch). Only add one when the API genuinely cannot confirm the write synchronously.
name is setadaptorType exact case matches connection type (request.yml > adaptorType)_connectionId references a valid, online connection (skip for AI imports without BYOK)http{} for HTTPImport, netsuite_da{} for NetSuiteDistributedImport, etc.)celigo http-connectors list) -- hand-written config only because no connector covers the app or operationhttp.method and http.relativeURI are set (http.yml)netsuite_da.operation and netsuite_da.recordType are set (netsuitedistributed.yml)rdbms.queryType is per_record or bulk_insert -- NOT legacy insert/update (rdbms.yml)salesforce.sObjectType and salesforce.operation are set (salesforce.yml)type matches the import's adaptorTypepageProcessors[] entry, not on the import itselfoneToMany: true and pathToMany is set to the child array pathmapping.fields[] / mapping.lists[], not mappings[]set command handles this.rest: block creates a legacy RESTImport. Use only http: for new imports.numIgnore on the job if records seem to bypass a step.{{blob}}. In a multipart/form-data parts array, the file part must be "type": "attachment" with "value": "{{blob}}" -- any other value 422s. Never hardcode the MIME boundary; the platform generates it. Fields the API wants alongside the file ride as "type": "inline" parts.bodyKey/blobKey in logs is an artifact, not a payload field. Audit and debug logs never show the assembled multipart body -- an internal storage pointer appears where the body would be. But if the destination actually received the literal string bodyKey or blobKey, the file part is misconfigured (an inline part where an attachment belongs, or a blobKeyPath that doesn't resolve).| Error | Cause | Fix |
|---|---|---|
422 adaptorType invalid | Wrong case | Use exact case from decision matrix: HTTPImport, NetSuiteDistributedImport, etc. |
422 _connectionId required | Missing connection | Set _connectionId to a valid connection ID |
422 queryType invalid | Legacy Snowflake value | Use per_record or bulk_insert, not insert/update |
422 distributed required | Missing NetSuite flag | Use NetSuiteDistributedImport with distributed: true on connection |
422 mapping invalid | Wrong mapper version | NetSuite/Salesforce use Mapper 1.0 (mapping.fields[]), not Mapper 2.0 (mappings[]) |
422 attachment value invalid | Multipart file part is not {{blob}} | Set the file part to "type": "attachment", "value": "{{blob}}"; let the platform generate the boundary |
# Search HTTP connectors
celigo http-connectors list | grep -i "<application-name>"
celigo http-connectors get <id> --full # see endpoints, resources, auth config
# Drill into the endpoints the connector defines for imports
celigo http-connectors catalog <id> --resource-type import --published-only
celigo http-connectors endpoint-detail <id> --resource-type import --resource-id <rid> --endpoint-id <epid>
# Search trading partner connectors (EDI, AS2)
celigo tp-connectors listceligo metadata types <connectionId> # List record types / sObjects / tables
celigo metadata fields <connectionId> <type> # List fields for an entity type# CRUD
celigo imports list
celigo imports get <id>
celigo imports create < import.json
celigo imports update <id> < import.json
celigo imports set <id> key=value [key2=value2 ...]
celigo imports delete <id> [-y]
# Invoke (test submission without creating a job)
echo '[{"name":"test"}]' | celigo imports invoke <id>
# Clone and connection management
echo '{"connectionMap":{"oldConnId":"newConnId"}}' | celigo imports clone <id>
celigo imports replace-connection <id> <newConnectionId>
# Discovery
celigo templates marketplace
celigo http-connectors list
celigo http-connectors catalog <id> --resource-type import --published-only
celigo http-connectors endpoint-detail <id> --resource-type import --resource-id <rid> --endpoint-id <epid>
celigo tp-connectors list
celigo metadata types <connectionId>
celigo metadata fields <connectionId> <entityType>
# Debug
celigo imports enable-debug <id> [--duration <minutes>]
celigo imports disable-debug <id>