Commit Graph

5444 Commits

Author SHA1 Message Date
Elixir Kook 7d0ff2dc85
YARN-10156. Fix typo 'complaint' which means quite different in Federation.md (#1856)
(cherry picked from commit d608e94f92)
2020-02-26 17:32:09 +09:00
Szilard Nemeth 92ad3bd099 YARN-10143. YARN-10101 broke Yarn logs CLI. Contributed by Adam Antal 2020-02-24 21:27:31 +01:00
Sunil G de63115a2a YARN-10139. ValidateAndGetSchedulerConfiguration API fails when cluster max allocation > default 8GB. Contributed by Prabhu Joseph.
(cherry picked from commit 6526f95bd2)
2020-02-19 11:18:19 +05:30
Szilard Nemeth 6aec712c6c YARN-10101. Support listing of aggregated logs for containers belonging to an application attempt. Contributed by Adam Antal 2020-02-11 09:18:44 +01:00
Sunil G 95b1cbcbd4 YARN-10109. Allow stop and convert from leaf to parent queue in a single Mutation API call. Contributed by Prabhu Joseph
(cherry picked from commit 28f730b317)
2020-02-09 21:15:31 +05:30
Jonathan Hung aca930402c YARN-10116. Expose diagnostics in RMAppManager summary
(cherry picked from commit 314e2f9d2e)
2020-02-05 11:17:03 -08:00
Prabhu Joseph 7136ebbb7a YARN-10022. Add RM Rest API to validate a CapacityScheduler Config with delta change
Contributed by Kinga Marton.

(cherry-picked from commit 1ab9c692fa)
2020-02-04 14:06:23 +05:30
Eric Badger 5736ecd123 YARN-10084. Allow inheritance of max app lifetime / default app lifetime. Contributed by Eric Payne. 2020-01-29 04:07:28 +00:00
Abhishek Modi be412546be YARN-9790. Failed to set default-application-lifetime if maximum-application-lifetime is less than or equal to zero. Contributed by kyungwan nam.
(cherry picked from commit d2d963f3d4)
2020-01-23 15:23:45 +00:00
Szilard Nemeth dcc16b07fa YARN-10083. Addendum to fix compilation error due to missing import 2020-01-22 17:18:03 +01:00
Szilard Nemeth 1e7679035f YARN-7913. Improve error handling when application recovery fails with exception. Contributed by Wilfred Spiegelenburg 2020-01-22 16:50:15 +01:00
Szilard Nemeth da416c826f YARN-10083. Provide utility to ask whether an application is in final status. Contributed by Adam Antal 2020-01-22 16:18:35 +01:00
Szilard Nemeth bbdc39c13e YARN-9462. TestResourceTrackerService.testNodeRemovalGracefully fails sporadically. Contributed by Prabhu Joseph 2020-01-22 15:49:35 +01:00
Szilard Nemeth c411a1468f YARN-8148. Update decimal values for queue capacities shown on queue status CLI. Contributed by Prabhu Joseph 2020-01-20 09:29:52 +01:00
Szilard Nemeth 805589fc71 YARN-9970. Refactor TestUserGroupMappingPlacementRule#verifyQueueMapping. Contributed by Manikandan R 2020-01-16 18:51:56 +01:00
Szilard Nemeth 0c2e312fef YARN-10028. Integrate the new abstract log servlet to the JobHistory server. Contributed by Adam Antal 2020-01-15 10:51:02 +01:00
Szilard Nemeth 6a7dfb3bf3 YARN-10026. Pull out common code pieces from ATS v1.5 and v2. Contributed by Adam Antal 2020-01-12 13:54:08 +01:00
Eric E Payne 3ba0fd1e50 YARN-9018. Add functionality to AuxiliaryLocalPathHandler to return all locations to read for a given path. Contributed by Kuhu Shukla (kshukla)
(cherry picked from commit 93233a7d6e)
2020-01-09 17:22:10 +00:00
Eric Badger 58db04ce15 YARN-8672. TestContainerManager#testLocalingResourceWhileContainerRunning occasionally times out. Contributed by Chandni Singh and Jim Brennan. 2020-01-08 19:44:43 +00:00
Eric E Payne 5e2be81fcb YARN-7387: org.apache.hadoop.yarn.server.resourcemanager.scheduler.capacity.TestIncreaseAllocationExpirer fails intermittently. Contributed by Jim Brennan (Jim_Brennan)
(cherry picked from commit b1e07d27cc)
2020-01-08 19:30:09 +00:00
Eric E Payne b20ce118da YARN-10072: TestCSAllocateCustomResource failures. Contributed by Jim Brennan (Jim_Brennan)
(cherry picked from commit 6899be5a17)
2020-01-08 17:42:41 +00:00
Prabhu Joseph ad98a30810 YARN-10053. Use Shared Group Mapping Service in Placement Rules. Contributed by Wilfred Spiegelenburg.
(cherry Picked from commit 217b56ffdd)
2020-01-02 14:30:01 +05:30
Akira Ajisaka 85da5cb870
YARN-10055. bower install fails. (#1778)
(cherry picked from commit 34ff7dbaf5)
2019-12-23 20:26:40 +09:00
Eric Badger 355ec33416 YARN-10009. In Capacity Scheduler, DRC can treat minimum user limit percent as a max when custom resource is defined. Contributed by Eric Payne 2019-12-20 19:32:36 +00:00
Jonathan Hung 0707d0a0ae YARN-9894. CapacitySchedulerPerf test for measuring hundreds of apps in a large number of queues. Contributed by Eric Payne
(cherry picked from commit 7b93575b92)
2019-12-18 13:18:22 -08:00
Jonathan Hung 01e2c996ff YARN-10039. Allow disabling app submission from REST endpoints
(cherry picked from commit fddc3d55c3)
2019-12-18 10:51:57 -08:00
Eric Badger 18b515322c YARN-10033. TestProportionalCapacityPreemptionPolicy not initializing vcores for effective max resources. Contributed by Eric Payne.
(cherry picked from commit f47dcf2d4c)
2019-12-17 17:36:52 +00:00
Akira Ajisaka a3cdba7713
YARN-9985. Unsupported transitionToObserver option displaying for rmadmin command. Contributed by Ayush Saxena.
(cherry picked from commit dc66de7448)
2019-12-09 18:38:47 +09:00
Jonathan Hung 9228e3f0ad YARN-10012. Guaranteed and max capacity queue metrics for custom resources. Contributed by Manikandan R
(cherry picked from commit 92bce918dc)
2019-12-08 16:36:08 -08:00
Jonathan Hung 8badcd989e Revert "YARN-10012. Guaranteed and max capacity queue metrics for custom resources"
This reverts commit b5235f1ed0.
2019-12-08 16:36:00 -08:00
Jonathan Hung b5235f1ed0 YARN-10012. Guaranteed and max capacity queue metrics for custom resources
(cherry picked from commit 92bce918dc)
2019-12-08 16:03:34 -08:00
Sunil G 69dc329acc YARN-4901. QueueMetrics needs to be cleared before MockRM is initialized. Contributed by Peter Bacsko.
(cherry picked from commit 002dcc4ebf)
2019-12-08 14:40:45 -08:00
Szilard Nemeth 2ad7b90505 YARN-9993. Remove incorrectly committed files from YARN-9011. Contributed by Wilfred Spiegelenburg 2019-11-28 12:39:54 +01:00
Sunil G f9b872b6ec YARN-9949. Add missing queue configs for root queue in RMWebService#CapacitySchedulerInfo.
Contributed by Prabhu Joseph.
2019-11-27 23:14:33 +05:30
Szilard Nemeth 8eda9fcab8 YARN-9937. Add missing queue configs in RMWebService#CapacitySchedulerQueueInfo.
Contributed by Prabhu Joseph.
2019-11-27 22:24:06 +05:30
Szilard Nemeth 3fc8930129 YARN-9011. Race condition during decommissioning. Contributed by Peter Bacsko 2019-11-26 14:26:58 +01:00
HUAN-PING SU 59a6261e81
YARN-9966. Code duplication in UserGroupMappingPlacementRule (#1709)
(cherry picked from commit f8e36e03b4)
2019-11-25 15:29:37 +09:00
Szilard Nemeth dcc453b4b8 YARN-9968. Public Localizer is exiting in NodeManager due to NullPointerException. Contributed by Tarun Parimi 2019-11-22 12:59:35 +01:00
Tao Yang af495192a5 YARN-9838. Fix resource inconsistency for queues when moving app with reserved container to another queue. Contributed by jiulongzhu. 2019-11-22 16:14:16 +08:00
Eric Yang 6951689f4c YARN-9983. Fixed typo in YARN Service overview.
Contributed by Denes Gerencser
2019-11-19 14:19:49 -05:00
Abhishek Modi 31591bb296 YARN-9791. Queue Mutation API does not allow to remove a config. Contributed by Prabhu Joseph.
(cherry picked from commit 751b5a1ac8)
2019-11-19 16:49:14 +05:30
Sunil G b04f152876 YARN-9909. Offline Format of YarnConfigurationStore. Contributed by Prabhu Joseph 2019-11-19 16:10:02 +05:30
Sunil G c1ec51696c YARN-8373. RM Received RMFatalEvent of type CRITICAL_THREAD_CRASH. Contributed by Wilfred Spiegelenburg.
(cherry picked from commit ea68756c0c)
2019-11-19 14:12:03 +05:30
Sunil G 049279bb66 YARN-9984. FSPreemptionThread can cause NullPointerException while app is unregistered with containers running on a node. Contributed by Wilfred Spiegelenburg.
(cherry picked from commit 215f2052fc)
2019-11-19 14:05:11 +05:30
Jonathan Eagles 254e18dcaf Revert "YARN-9949. Add missing queue configs for root queue in RMWebService#CapacitySchedulerInfo. Contributed by Prabhu Joseph."
This reverts commit 11c763c220.
2019-11-05 15:10:01 -06:00
Sunil G 597b315811 YARN-9950. Unset Ordering Policy of Leaf/Parent queue converted from Parent/Leaf queue respectively. Contributed by Prabhu Joseph.
(cherry picked from commit 51e7d1b37e)
2019-11-04 23:28:39 +05:30
Sunil G 11c763c220 YARN-9949. Add missing queue configs for root queue in RMWebService#CapacitySchedulerInfo. Contributed by Prabhu Joseph.
(cherry picked from commit d462308e04)
2019-11-03 08:48:04 +05:30
Jonathan Hung 5d2ffcc7aa Make upstream aware of 2.10.0 release
(cherry picked from commit 7663db59c097c82eeed2df7a91168a4d7123c96b)
2019-10-30 20:59:20 -07:00
Eric Badger fa6b27ea8d YARN-9914. Use separate configs for free disk space checking for full and not-full disks. Contributed by Jim Brennan
(cherry picked from commit eef34f2d87)
2019-10-25 17:15:48 +00:00
Eric E Payne ea574087d1 YARN-9915: Fix FindBug issue in QueueMetrics. Contributed by Prabhu Joseph.
(cherry picked from commit 83d148074f)
2019-10-21 20:56:40 +00:00
Eric E Payne 23b72d8ae1 YARN-9773: Add QueueMetrics for Custom Resources. Contributed by Manikandan R.
(cherry picked from commit a5034c7988)
2019-10-16 21:13:02 +00:00
Sunil G 9672b81fa3 YARN-9900. Revert to previous state when Invalid Config is applied and Refresh Support in SchedulerConfig Format. Contributed by Prabhu Joseph.
(cherry picked from commit 090f73a9aa)
2019-10-16 18:15:34 +05:30
Haibo Chen 3a5474c61e YARN-8842. Expose metrics for custom resource types in QueueMetrics. (Contributed by Szilard Nemeth)
(cherry picked from commit 84e22a6af4)
2019-10-15 22:14:33 +00:00
Haibo Chen 1344823d4d YARN-8750. Refactor TestQueueMetrics. (Contributed by Szilard Nemeth)
(cherry picked from commit e60b797c88)
2019-10-15 15:32:01 +00:00
Szilard Nemeth b10fdd136a YARN-8453. Additional Unit tests to verify queue limit and max-limit with multiple resource types. Contributed by Adam Antal 2019-10-15 13:24:59 +02:00
Akira Ajisaka eb4bd54938
YARN-7243. Moving logging APIs over to slf4j in hadoop-yarn-server-resourcemanager. (#1634)
Signed-off-by: Akira Ajisaka <aajisaka@apache.org>
(cherry picked from commit e40e2d6ad5)

 Conflicts:
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/main/java/org/apache/hadoop/yarn/server/resourcemanager/rmnode/RMNodeImpl.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/main/java/org/apache/hadoop/yarn/server/resourcemanager/scheduler/AbstractYarnScheduler.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/main/java/org/apache/hadoop/yarn/server/resourcemanager/scheduler/capacity/CapacityScheduler.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/main/java/org/apache/hadoop/yarn/server/resourcemanager/scheduler/capacity/LeafQueue.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/main/java/org/apache/hadoop/yarn/server/resourcemanager/volume/csi/VolumeManagerImpl.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/main/java/org/apache/hadoop/yarn/server/resourcemanager/volume/csi/lifecycle/VolumeImpl.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/test/java/org/apache/hadoop/yarn/server/resourcemanager/ApplicationMasterServiceTestBase.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/test/java/org/apache/hadoop/yarn/server/resourcemanager/TestApplicationMasterLauncher.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/test/java/org/apache/hadoop/yarn/server/resourcemanager/TestApplicationMasterServiceInterceptor.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/test/java/org/apache/hadoop/yarn/server/resourcemanager/TestResourceManager.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/test/java/org/apache/hadoop/yarn/server/resourcemanager/rmapp/TestRMAppTransitions.java
	hadoop-yarn-project/hadoop-yarn/hadoop-yarn-server/hadoop-yarn-server-resourcemanager/src/test/java/org/apache/hadoop/yarn/server/resourcemanager/scheduler/fair/TestFairSchedulerConfiguration.java
2019-10-10 15:27:54 +09:00
Szilard Nemeth da35a22083 Revert "YARN-9128. Use SerializationUtils from apache commons to serialize / deserialize ResourceMappings. Contributed by Zoltan Siegl"
This reverts commit 42177e8b78.
2019-10-09 19:58:46 +02:00
Szilard Nemeth 05966ce204 YARN-9552. FairScheduler: NODE_UPDATE can cause NoSuchElementException. Contributed by Peter Bacsko. 2019-10-09 14:18:06 +02:00
Szilard Nemeth 42177e8b78 YARN-9128. Use SerializationUtils from apache commons to serialize / deserialize ResourceMappings. Contributed by Zoltan Siegl
(cherry picked from commit 6f1ab95168)
2019-10-09 13:28:01 +02:00
Szilard Nemeth 0ddb48a303 YARN-9356. Add more tests to ratio method in TestResourceCalculator. Contributed by Zoltan Siegl
(cherry picked from commit 35f093f5b3)
2019-10-09 13:12:03 +02:00
Sunil G 9cb3ab058f YARN-9873. Mutation API Config Change need to update Version Number. Contributed by Prabhu Joseph
(cherry picked from commit be901f4962)
2019-10-09 15:54:09 +05:30
Jonathan Hung ad3c98456d YARN-9760. Support configuring application priorities on a workflow level. Contributed by Varun Saxena
(cherry picked from commit eebd313d76ed742fe82292bd8c0184970cdc5692)
2019-10-08 11:17:05 -07:00
Szilard Nemeth 9b4aba49d1 YARN-6715. Fix documentation about NodeHealthScriptRunner. Contributed by Peter Bacsko 2019-10-08 17:41:46 +02:00
Sunil G 5704f15589 Revert "YARN-9873. Mutation API Config Change updates Version Number. Contributed by Prabhu Joseph"
This reverts commit 3a0afcfb7f.
2019-10-05 09:16:04 +05:30
Sunil G 3a0afcfb7f YARN-9873. Mutation API Config Change updates Version Number. Contributed by Prabhu Joseph
(cherry picked from commit 4510970e2f)
2019-10-04 21:49:49 +05:30
Sunil G 312cfa994a YARN-9801. SchedConfCli does not work wiwith https mode. Contributed by Prabhu Joseph
(cherry picked from commit 99cd7572f1)
2019-10-01 20:07:18 +05:30
bibinchundatt 5cd6eb2a18 YARN-9858. Optimize RMContext getExclusiveEnforcedPartitions. Contributed by Jonathan Hung. 2019-10-01 16:04:58 +05:30
Sunil G 52f815d39d YARN-9864. Format CS Configuration present in Configuration Store. Contributeed by Prabhu Joseph
(cherry picked from commit 137546a78a)
2019-10-01 09:09:22 +05:30
Jonathan Hung f4f210d2e5 Addendum to YARN-9730. Support forcing configured partitions to be exclusive based on app node label
(cherry picked from commit d86a1acc866cbda845fb3896dc824baf12217383)
2019-09-25 17:49:37 -07:00
Jonathan Hung 806c7b7dfb YARN-9730. Support forcing configured partitions to be exclusive based on app node label
(cherry picked from commit 73a044a63822303f792183244e25432528ecfb1e)
2019-09-24 13:51:54 -07:00
Jonathan Hung a1fa9a8a7f YARN-9762. Add submission context label to audit logs. Contributed by Manoj Kumar
(cherry picked from commit 3d78b1223d)
2019-09-23 13:12:57 -07:00
Sunil G 3e0025d877 YARN-9833. Race condition when DirectoryCollection.checkDirs() runs during container launch. Contributed by Peter Bacsko.
(cherry picked from commit c474e24c0b)
2019-09-18 09:22:48 +05:30
Weiwei Yang 7ec229244a YARN-2255. YARN Audit logging not added to log4j.properties. Contributed by Aihua Xu.
(cherry picked from commit f8c14326ee)
2019-09-18 09:19:38 +08:00
Eric Yang 345ef049df YARN-9837. Fixed reading YARN Service JSON spec file larger than 128k.
Contributed by Tarun Parimi

(cherry picked from commit eefe9bc85c)
2019-09-17 13:22:53 -04:00
Jonathan Hung 1dbf87c9ff YARN-9824. Fall back to configured queue ordering policy class name
(cherry picked from commit f8f8598ea5)
2019-09-10 15:26:57 -07:00
bibinchundatt e10050678d YARN-8948. PlacementRule interface should be for all YarnSchedulers. Contributed by Bibin A Chundatt.
(cherry picked from commit a68d766e87)
2019-09-09 19:03:40 -07:00
Abhishek Modi f6cc887f35 YARN-9821. NM hangs at serviceStop when ATSV2 Backend Hbase is Down. Contributed by Prabhu Joseph. 2019-09-09 15:44:45 +05:30
Jonathan Hung 1f0449ddfb YARN-9820. RM logs InvalidStateTransitionException when app is submitted. Contributed by Prabhu Joseph 2019-09-09 00:24:17 -07:00
Jonathan Hung 45220d1157 YARN-9764. Print application submission context label in application summary. Contributed by Manoj Kumar
(cherry picked from commit 43e389b980)
2019-09-08 19:11:47 -07:00
Rohith Sharma K S 7d5bb2ebb7 Preparing for 3.2.2-SNAPSHOT development. 2019-09-07 08:52:08 +05:30
Wangda Tan 0e77347972 YARN-9813. RM does not start on JDK11 when UIv2 is enabled. (Adam Antal/Eric Yang via wangda)
Change-Id: I18b8edc930b2efa0652f59c246931ad0d46827f3
(cherry picked from commit 34b82e6da0)
2019-09-06 19:19:05 -07:00
Tao Yang 9ee257e353 YARN-8995. Log events info in AsyncDispatcher when event queue size cumulatively reaches a certain number every time(addendum). Contributed by Jonathan Hung. 2019-09-07 08:25:15 +08:00
Tao Yang cfce39023d YARN-9795. ClusterMetrics to include AM allocation delay. Contributed by Fengnan Li. 2019-09-07 07:53:46 +08:00
Jonathan Hung 9c9ff07249 YARN-9763. Print application tags in application summary. Contributed by Manoj Kumar 2019-09-06 10:49:43 -07:00
Jonathan Hung 1f685efc73 YARN-9761. Allow overriding application submissions based on server side configs. Contributed by Pralabh Kumar 2019-09-06 10:13:12 -07:00
Jonathan Hung 79ca399a30 YARN-9810. Add queue capacity/maxcapacity percentage metrics. Contributed by Shubham Gupta
(cherry picked from commit 0ccf4b0fe1)
2019-09-05 14:05:45 -07:00
Billie Rinaldi 66627749d0 YARN-9718. Fixed yarn.service.am.java.opts shell injection. Contributed by Eric Yang
(cherry picked from commit 2e2e5401f2)
2019-09-05 12:54:20 -07:00
Rohith Sharma K S 4d9c5300e2 YARN-8567. Fetching yarn logs fails for long running application if it is not present in timeline store. Contributed by Tarun Parimi. 2019-09-05 18:16:35 +05:30
Eric Yang b87a727ff4 YARN-9374. Improve Timeline service resilience when HBase is unavailable.
Contributed by Prabhu Joseph and Szilard Nemeth
2019-09-05 16:32:18 +05:30
Eric Yang 02779cdc3a YARN-8499 ATSv2 Generalize TimelineStorageMonitor.
Contributed by Prabhu Joseph
2019-09-05 16:31:35 +05:30
Eric Yang 6110af2d1d YARN-7537. Add ability to load hbase config from distributed file system.
Contributed by Prabhu Joseph
2019-09-05 16:28:04 +05:30
Vrushali C bcacb57114 YARN-9335 [atsv2] Restrict the number of elements held in timeline collector when backend is unreachable for async calls. Contributed by Abhishesk Modi. 2019-09-05 16:27:39 +05:30
Vrushali C 6acc1a2bd0 YARN-9382 Publish container killed, paused and resumed events to ATSv2. Contributed by Abhishesk Modi. 2019-09-05 15:39:38 +05:30
Vrushali C f52a88fdc8 YARN-9303 Username splits won't help timelineservice.app_flow table. Contributed by Prabhu Joseph. 2019-09-05 15:39:38 +05:30
Giovanni Matteo Fumarola 998aa3de2c YARN-9418. ATSV2 /apps//entities/YARN_CONTAINER rest api does not show metrics. Contributed by Prabhu Joseph. 2019-09-05 15:39:38 +05:30
Rohith Sharma K S 8de93fca3c YARN-9389. FlowActivity and FlowRun table prefix is wrong. Contributed by Prabhu Joseph. 2019-09-05 15:39:38 +05:30
Rohith Sharma K S 0ccc5a2695 YARN-9387. Update document for ATS HBase Custom tablenames (-entityTableName). Contributed by Prabhu Joseph. 2019-09-05 15:39:38 +05:30
Vrushali C d451ff7534 YARN-3841 [atsv2 Storage implementation] Adding retry semantics to HDFS backing storage. Contributed by Abhishek Modi. 2019-09-05 15:39:38 +05:30
Vrushali C 66e1599761 YARN-3879 [Storage implementation] Create HDFS backing storage implementation for ATS reads. Contributed by Abhishek Modi. 2019-09-05 15:39:38 +05:30
Tao Yang 6f9764076a YARN-8995. Log events info in AsyncDispatcher when event queue size cumulatively reaches a certain number every time. Contributed by zhuqi. 2019-09-05 16:53:16 +08:00