[ https://issues.apache.org/jira/browse/BEAM-991?page=com.atlassian.jira.plugin.system.issuetabpanels:comment-tabpanel&focusedCommentId=15702845#comment-15702845 ]
Daniel Halperin commented on BEAM-991: -------------------------------------- Hi [~JoshuaFox], [~vikasrk], I agree with Vikas, though I would run some performance experiments with complex types (proto3 map type with some complex values filled in) to confirm there's no perf impact ;) Joshua, do you know whether you'll be able to take this on? Thanks, Dan > DatastoreIO Write should flush early for large batches > ------------------------------------------------------ > > Key: BEAM-991 > URL: https://issues.apache.org/jira/browse/BEAM-991 > Project: Beam > Issue Type: Bug > Components: sdk-java-gcp > Reporter: Vikas Kedigehalli > Assignee: Vikas Kedigehalli > > If entities are large (avg size > 20KB) then the a single batched write (500 > entities) would exceed the Datastore size limit of a single request (10MB) > from https://cloud.google.com/datastore/docs/concepts/limits. > First reported in: > http://stackoverflow.com/questions/40156400/why-does-dataflow-erratically-fail-in-datastore-access -- This message was sent by Atlassian JIRA (v6.3.4#6332)