Chiebot-Mirror/grpc - grpc - Gitea: Git with a cup of tea

Commit Graph

Author	SHA1	Message	Date
Vignesh Babu	9c59671936	[testing] Skip more flaky event engine tests (#33160 ) Expand the set with more new flaky tests.	2 years ago
Luwei Ge	05d5a04186	[Audit Logging] Custom audit logger parsing in xDS registry. (#32970 )	2 years ago
Craig Tiller	5fac4ad47b	[fuzzing] Improve OSA distance performance (#33149 ) Early out evaluating this function where we can, and use macros to eliminate function calls in debug builds. Takes per-example time from 5400ms to 1200ms in debug asan builds. <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. --> --------- Co-authored-by: ctiller <ctiller@users.noreply.github.com>	2 years ago
AJ Heller	b35ea15265	[reland][fuzzing] Client channel resolver fuzzer (#33153 ) This reverts commit `1624542ea4`, relanding https://github.com/grpc/grpc/pull/32956 Because of some proto dependency and build problems internally, I've removed the ServiceConfig proto fuzzing component. These build issues can hopefully be resolved soon, and then we can re-add the deleted implementation from commit [`b078c9c`](`b078c9c015`) in this PR.	2 years ago
Mark D. Roth	a78001a087	[resolver] remove unused ctor for ServerAddress (#33148 ) Co-authored-by: markdroth <markdroth@users.noreply.github.com>	2 years ago
Vignesh Babu	63ecc4ba3e	[testing] Temporarily skip flaky event engine tests. (#33136 ) Based on flaky tests reported by dashboard: https://dashboards.corp.google.com/stubby_team.grpc_flaky_dashboard#1318s66d4f	2 years ago
AJ Heller	1624542ea4	Revert "[fuzzing] Client channel resolver fuzzer" (#33152 ) Reverts grpc/grpc#32956. Requires a cherrypick.	2 years ago
AJ Heller	252ebad341	[infra] Fix absl::Mutex check and remove all uses (#33144 ) `tools/run_tests/sanity/check_absl_mutex.sh` was broken, a missing paren crashed the script if run locally. It's unclear yet how our sanity checks were not complaining about this, `run_tests.py` does not save the log.	2 years ago
Craig Tiller	239d3e6857	[fuzzing] Allow core_end2end_test_fuzzer, api_fuzzer to change experiments (#33147 ) They were intended to be able to, but since these are currently frozen across the process it wasn't possible. Fix that for these fuzzers.	2 years ago
Craig Tiller	74ec5d1684	[promises] Improve logging, fix a rare bug (#33139 ) Rare bug: server initial metadata gets stranded in the outbound pipe. (fix is a little unpleasant, but we'll do better at the five pipes stage) <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago
AJ Heller	030ecf60ec	[fuzzing] Client channel resolver fuzzer (#32956 ) Co-authored-by: drfloob <drfloob@users.noreply.github.com>	2 years ago
Craig Tiller	dcef4bb981	[party] Disable mutex test on mac (#33143 ) This has been flaky, but we don't need this scenario tested right now (especially on Mac), so disable that part of the testing.	2 years ago
Luwei Ge	6df358cf6a	[Audit Logging] Stdout logger implementation (#33026 ) The logger uses `absl::FPrintF` to write to stdout. After reading a number of sources online, I got the impression that `std::fwrite` which is used by `absl::FPrintF` is atomic so there is no locking required here. --------- Co-authored-by: rockspore <rockspore@users.noreply.github.com>	2 years ago
Craig Tiller	cd44a2433e	[call] Dont take grpclb_client_stats from the app (#33118 ) This metadata doesn't actually encode so passing it through from an app will force a crash. <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago
Vignesh Babu	915d7c4a70	[Fuzzing] Bound RunAfter duration in fuzzing event engine (#33128 ) Bounds duration to 1 year. Fixes b/258949216	2 years ago
Yijie Ma	3526defc19	[JsonWriter] Do not break in EscapeString when encountering a null byte (#33127 ) Instead just Utf-16 encode the null byte when dumping the value to a string form. <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago
Craig Tiller	997af8d073	[api_fuzzer] Attempt to clean up fuzzer memory leak (#33120 ) <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago
Mark D. Roth	59dbdfeea2	[xds_client_fuzzer] fail bootstrap parsing if xds_servers is empty (#33119 ) b/269022924	2 years ago
Mark D. Roth	5ae1cfcce3	[xds_client_fuzzer] fix null pointer dereference in `FakeXdsTransport::TriggerConnectionFailure()` (#33117 ) b/259358608	2 years ago
Mark D. Roth	13133ae703	[xds_client_fuzzer] fix bug in fake transport (#33115 ) Fixes `FakeXdsTransport` to remove itself from the map in `FakeXdsTransportFactory` when it gets orphaned by the `XdsClient`, so that a subsequent creation of a new transport for the same server does not trigger an assertion due to the transport already existing in the map. Fixes internal b/259362837.	2 years ago
Craig Tiller	66d9f52fbd	[api-fuzzer] Fix memory leak (#33109 ) ApiFuzzer::CreateChannel() called twice creates two channels but doesn't delete the first. Choose some reasonable behavior.	2 years ago
Craig Tiller	123811399b	[promises] Remove bad log statement (#33113 ) Was leading to a nullptr deref, and we just don't need this one anymore. <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago
Craig Tiller	18d369a6f4	[fuzzing] Avoid initialization order fiasco in core_end2end_test_fuzzer (#33108 )	2 years ago
Craig Tiller	ee0cf2fada	[filter-fuzzer] Delete this fuzzer until I can spend time on it (#33096 ) It's not finished and won't be for a bit...	2 years ago
Craig Tiller	9760ce9d0a	[end2end] Shorten corpora filenames (#33095 ) Avoids long path name problems on Windows <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago
Mark D. Roth	8fdfb22848	[JSON] generalize handling of RefCountedPtr<> (#33048 ) Also remove a check in the weighted_target LB policy that I somehow missed in #32932.	2 years ago
Yash Tibrewal	1b01336504	[Logging] Fix flakiness in test (#33077 ) Change - just make sure that before we verify the logging entries, we'll wait for the expected number of entries to show up. Logging_test has been recently reported as flaky. Sample failure - https://source.cloud.google.com/results/invocations/ba581ad8-b652-4b9d-af56-07593f5d2deb/targets/%2F%2Ftest%2Fcpp%2Fext%2Ffilters%2Flogging:logging_test@poller%3Depoll1/tests Failed to reproduce, but i have a suspicion that what's happening is that the logging for the half-close events on the server side can happen after the call has ended.. It is currently being logged after the server-trailer (which has the status).	2 years ago
Craig Tiller	4674f2ccf7	[fuzz] Turn core end2end tests into fuzzers (#33013 ) Add a new binary that runs all core end2end tests in fuzzing mode. In this mode FuzzingEventEngine is substituted for the default event engine. This means that time is simulated, as is IO. The FEE gets control of callback delays also. In our tests the `Step()` function becomes, instead of a single call to `completion_queue_next`, a series of calls to that function and `FuzzingEventEngine::Tick`, driving forward the event loop until progress can be made. PR guide: --- New binaries `core_end2end_test_fuzzer` - the new fuzzer itself `seed_end2end_corpus` - a tool that produces an interesting seed corpus Config changes for safe fuzzing The implementation tries to use the config fuzzing work we've previously deployed in api_fuzzer to fuzz across experiments. Since some experiments are far too experimental to be safe in such fuzzing (and this will always be the case): - a new flag is added to experiments to opt-out of this fuzzing - a new hook is added to the config system to allow variables to re-write their inputs before setting them during the fuzz Event manager/IO changes Changes are made to the event engine shims so that tcp_server_posix can run with a non-FD carrying EventEngine. These are in my mind a bit clunky, but they work and they're in code that we expect to delete in the medium term, so I think overall the approach is good. Changes to time A small tweak is made to fix a bug initializing time for fuzzers in time.cc - we were previously failing to initialize `g_process_epoch_cycles` Changes to `Crash` A version that prints to stdio is added so that we can reliably print a crash from the fuzzer. Changes to CqVerifier Hooks are added to allow the top level loop to hook the verification functions with a function that steps time between CQ polls. Changes to end2end fixtures State machinery moves from the fixture to the test infra, to keep the customizations for fuzzing or not in one place. This means that fixtures are now just client/server factories, which is overall nice. It did necessitate moving some bespoke machinery into h2_ssl_cert_test.cc - this file is beginning to be problematic in borrowing parts but not all of the e2e test machinery. Some future PR needs to solve this. A cq arg is added to the Make functions since the cq is now owned by the test and not the fixture. Changes to test registration `TEST_P` is replaced by `CORE_END2END_TEST` and our own test registry is used as a first depot for test information. The gtest version of these tests: queries that registry to manually register tests with gtest. This ultimately changes the name of our tests again (I think for the last time) - the new names are shorter and more readable, so I don't count this as a regression. The fuzzer version of these tests: constructs a database of fuzzable tests that it can consult to look up a particular suite/test/config combination specified by the fuzzer to fuzz against. This gives us a single fuzzer that can test all 3k-ish fuzzing ready tests and cross polinate configuration between them. Changes to test config The zero size registry stuff was causing some problems with the event engine feature macros, so instead I've removed those and used GTEST_SKIP in the problematic tests. I think that's the approach we move towards in the future. Which tests are included Configs that are compatible - those that do not do fd manipulation directly (these are incompatible with FuzzingEventEngine), and those that do not join threads on their shutdown path (as these are incompatible with our cq wait methodology). Each we can talk about in the future - fd manipulation would be a significant expansion of FuzzingEventEngine, and is probably not worth it, however many uses of background threads now should probably evolve to be EventEngine::Run calls in the future, and then would be trivially enabled in the fuzzers. Some tests currently fail in the fuzzing environment, a `SKIP_IF_FUZZING` macro is used for these few to disable them if in the fuzzing environment. We'll burn these down in the future. Changes to fuzzing_event_engine Changes are made to time: an exponential sweep forward is used now - this catches small time precision things early, but makes decade long timers (we have them) able to be used right now. In the future we'll just skip time forward to the next scheduled timer, but that approach doesn't yet work due to legacy timer system interactions. Changes to port assignment: we ensure that ports are legal numbers before assigning them via `grpc_pick_port_or_die`. A race condition between time checking and io is fixed. --------- Co-authored-by: ctiller <ctiller@users.noreply.github.com>	2 years ago
Hannah Shi	ad2a5dd355	[ObjC] Cf event engine client (#33034 ) Added `//:gpr_platform` to cf_engine_test to fix build_cleaner check in the previous merge. More details in https://github.com/grpc/grpc/pull/33027	2 years ago
Mark D. Roth	2c423d277c	[outlier detection] fix crash with pick_first and add tests (#33069 ) Fixes #32967. Also fix incorrect defaults for `enforcementPercentage` fields.	2 years ago
AJ Heller	0ed3bb7955	[EventEngine] Disable more thread pool tests for legacy implementation (#33068 )	2 years ago
AJ Heller	63ec566f3e	[EventEngine] Reduce the size of some thread pool tests (#33055 ) The `CanStartLotsOfClosures` test was sometimes taking over 60s to run ([example](https://source.cloud.google.com/results/invocations/d96c89f9-03f9-43fd-a729-d744d7499532/targets;query=thread_pool_test/%2F%2Ftest%2Fcore%2Fevent_engine:thread_pool_test@poller%3Depoll1/log)). More often than not, though, the test would take < 5s ([example](https://source.cloud.google.com/results/invocations/95d32b32-5df7-4dd4-a82c-1024869b09c8/targets;query=thread_pool_test/%2F%2Ftest%2Fcore%2Fevent_engine:thread_pool_test/log)). Both examples are from before the tests changed with the introduction of the work-stealing thread pool (`3fb738b9b1`). This PR reduces the closure count to 500k for the `CanStartLotsOfClosures` test, and changes the blocking-closure scale-test to exercise the work stealing implementation alone.	2 years ago
Mark D. Roth	17315823c2	[client channel] assume LB policies start in CONNECTING state (#33009 ) Currently, we are not very consistent in what we assume the initial state of an LB policy will be and whether or not we assume that it will immediately report a new picker when it gets its initial address update; different parts of our code make different assumptions. This PR establishes the convention that LB policies will be assumed to start in state CONNECTING and will not be assumed to report a new picker immediately upon getting their initial address update, and we now assume that convention everywhere consistently. This is a preparatory step for changing policies like round_robin to delegate to pick_first, which I'm working on in #32692. As part of that change, we need pick_first to not report a connectivity state until it actually sees the connectivity state of the underlying subchannels, so that round_robin knows when to swap over to a new child list without reintroducing the problem fixed in #31939.	2 years ago
Esun Kim	37e9903ecb	[Build] Fix json error (#33051 ) To fix this error ``` test/core/security/grpc_authorization_engine_test.cc:88:32: error: unknown type name 'Json'; did you mean 'experimental::Json'? ParseAuditLoggerConfig(const Json&) override { ^~~~ experimental::Json ```	2 years ago
Mark D. Roth	1432fe4e4c	[JSON] make API public but experimental (#32987 ) This makes the JSON API visible as part of the C-core API, but in the `experimental` namespace. It will be used as part of various experimental APIs that we will be introducing in the near future, such as the audit logging API.	2 years ago
Ming-Chuan	6c2f4371bb	[Binder Transport] Flush ExecCtx in e2e test (#32971 ) WireWriter implementation schedules actions to be run by `ExecCtx`. We should flush pending actions before destructing `end2end_testing::g_transaction_processor`, which need to be alive to handle the scheduled actions. Otherwise, we get heap-use-after-free error because the testing fixture (`end2end_testing::g_transaction_processor`) is destructed before all the scheduled actions are run. This lowers end2end binder transport test failure rate from 0.23% to 0.15%, according to internal tool that runs the test for 15000 times under various configuration.	2 years ago
Mark D. Roth	e872fb91d9	[WRR] fix some edge cases in scheduler logic (#33045 ) This corresponds to two recent changes made to our internal implementation. See b/276292666 for details.	2 years ago
AJ Heller	3fb738b9b1	[EventEngine] Implement work-stealing in the EventEngine ThreadPool (#32869 ) This PR implements a work-stealing thread pool for use inside EventEngine implementations. Because of historical risks here, I've guarded the new implementation behind an experiment flag: `GRPC_EXPERIMENTS=work_stealing`. Current default behavior is the original thread pool implementation. Benchmarks look very promising: ``` bazel test \ --test_timeout=300 \ --config=opt -c opt \ --test_output=streamed \ --test_arg='--benchmark_format=csv' \ --test_arg='--benchmark_min_time=0.15' \ --test_arg='--benchmark_filter=_FanOut' \ --test_arg='--benchmark_repetitions=15' \ --test_arg='--benchmark_report_aggregates_only=true' \ test/cpp/microbenchmarks:bm_thread_pool ``` 2023-05-04: `bm_thread_pool` benchmark results on my local machine (64 core ThreadRipper PRO 3995WX, 256GB memory), comparing this PR to master: ![image](https://user-images.githubusercontent.com/295906/236315252-35ed237e-7626-486c-acfa-71a36f783d22.png) 2023-05-04: `bm_thread_pool` benchmark results in the Linux RBE environment (unsure of machine configuration, likely small), comparing this PR to master. ![image](https://user-images.githubusercontent.com/295906/236317164-2c5acbeb-fdac-4737-9b2d-4df9c41cb825.png) --------- Co-authored-by: drfloob <drfloob@users.noreply.github.com>	2 years ago
Yijie Ma	7df0e11755	[EventEngine] Change TXT lookup result type to std::vector<std::string> (#33030 ) One TXT lookup query can return multiple TXT records (see the following example). `EventEngine::DNSResolver` should return all of them to let the caller (e.g. `event_engine_client_channel_resolver`) decide which one they would use. ``` $ dig TXT wikipedia.org ; <<>> DiG 9.18.12-1+build1-Debian <<>> TXT wikipedia.org ;; global options: +cmd ;; Got answer: ;; ->>HEADER<<- opcode: QUERY, status: NOERROR, id: 49626 ;; flags: qr rd ra; QUERY: 1, ANSWER: 3, AUTHORITY: 0, ADDITIONAL: 1 ;; OPT PSEUDOSECTION: ; EDNS: version: 0, flags:; udp: 512 ;; QUESTION SECTION: ;wikipedia.org. IN TXT ;; ANSWER SECTION: wikipedia.org. 600 IN TXT "google-site-verification=AMHkgs-4ViEvIJf5znZle-BSE2EPNFqM1nDJGRyn2qk" wikipedia.org. 600 IN TXT "yandex-verification: 35c08d23099dc863" wikipedia.org. 600 IN TXT "v=spf1 include:wikimedia.org ~all" ``` Note that this change also deviates us from the iomgr's DNSResolver API which uses std::string as the result type. <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago
AJ Heller	ee0aaacbde	Revert "[ObjC] CF Stream Event Engine Client" (#33027 ) Reverts grpc/grpc#32924. This breaks the build again, unfortunately. From `test/core/event_engine/cf:cf_engine_test`: ``` error: module .../grpc/test/core/event_engine/cf:cf_engine_test does not depend on a module exporting 'grpc/support/port_platform.h' ``` @sampajano I recommend looking into CI tests to catch iOS problems before merging. We can enable EventEngine experiments in the CI generally once this PR lands, but this broken test is not one of those experiments. A normal build should have caught this. cc @HannahShiSFB	2 years ago
Hannah Shi	d0c1809840	[ObjC] CF Stream Event Engine Client (#32924 ) bazel build --config=macos --genrule_strategy=local --copt="-DGRPC_CFSTREAM=1" //test/cpp/end2end:cfstream_test succeeds Fixing failure described here: https://github.com/grpc/grpc/pull/32882#issuecomment-1512210309	2 years ago
Mark D. Roth	1fcaccdf5f	[client channel] Second attempt: use ChunkedVector for call attributes (#33015 ) Original was #33002, reverted in #33014. The second commit here adds a build visibility tag necessary to fix the internal build problems.	2 years ago
Esun Kim	303e568f27	[Build] Removed gRPC_PROTOBUF_PACKAGE_TYPE, supporting config only (#32988 ) `FindProtobuf` isn't working as Protobuf began to use Abseil so gRPC is now using `CONFIG` mode for protobuf module (Context: https://gitlab.kitware.com/cmake/cmake/-/issues/24321)	2 years ago
AJ Heller	18aab6ffb5	Revert "[client channel] use ChunkedVector for call attributes" (#33014 ) Reverts grpc/grpc#33002. Breaks internal builds: `.../privacy_context:filters does not depend on a module exporting '.../src/core/lib/channel/context.h'`	2 years ago
Mark D. Roth	2f89fd5528	[client channel] use ChunkedVector for call attributes (#33002 ) Change call attributes to be stored in a `ChunkedVector` instead of `std::map<>`, so that the storage can be allocated on the arena. This means that we're now doing a linear search instead of a map lookup, but the total number of attributes is expected to be low enough that that should be okay. Also, we now hide the actual data structure inside of the `ServiceConfigCallData` object, which required some changes to the `ConfigSelector` API. Previously, the `ConfigSelector` would return a `CallConfig` struct, and the client channel would then use the data in that struct to populate the `ServiceConfigCallData`. This PR changes that such that the client channel creates the `ServiceConfigCallData` before invoking the `ConfigSelector`, and it passes the `ServiceConfigCallData` into the `ConfigSelector` so that the `ConfigSelector` can populate it directly.	2 years ago
Luwei Ge	4c7da485c5	[xDS] Protect RBAC audit logging options field with environment variable. (#33004 ) The protection is added at `xds_http_rbac_filter.cc` where we read the new field. With this disabling the feature, nothing from things like `xds_audit_logger_registry.cc` shall be invoked.	2 years ago
Craig Tiller	ad41fe96b6	[promises] Re-enable C++ end2end tests (with fixes) (#32837 ) Makes some awkward fixes to compression filter, call, connected channel to hold the semantics we have upheld now in tests. Once the fixes described here https://github.com/grpc/grpc/blob/master/src/core/lib/channel/connected_channel.cc#L636 are in this gets a lot less ad-hoc, but that's likely going to be post-landing promises client & server side. We specifically need special handling for server side cancellation in response to reads wrt the inproc transport - which doesn't track cancellation thoroughly enough itself. <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. --> --------- Co-authored-by: ctiller <ctiller@users.noreply.github.com>	2 years ago
Craig Tiller	65a2a895af	[chttp2] Fix some fuzzer found bugs. (#33005 ) <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago
Luwei Ge	abc82b9e19	[Audit Logging] Audit logging support in authorization engines. (#32995 ) 1. `GrpcAuthorizationEngine` creates the logger from the given config in its ctor. 2. `Evaluate()` invokes audit logging when needed. --------- Co-authored-by: rockspore <rockspore@users.noreply.github.com>	2 years ago
Craig Tiller	79e46a6022	[channelz] Save some memory per channel (#32996 ) Whilst the per cpu counters probably help single channel contention, we think it's likely that they're a pessimization when taken fleetwide. <!-- If you know who should review your pull request, please assign it to that person, otherwise the pull request would get assigned randomly. If your pull request is for a specific language, please add the appropriate lang label. -->	2 years ago

... 3 4 5 6 7 ...

13656 Commits (60c1701f87cacf359aa1ad785728549eeef1a4b0)