Problem solution · C++

Groups of Strings

Groups of Strings: a C++ solution using disjoint set union. Learn the idea, check the complexity, and read the full code, with credit to walkccc LeetCode Solutions.

Technique
Disjoint set union
Source
walkccc LeetCode Solutions
Length
80 lines
Start with the idea.

Try the problem first. If you get stuck, read the approach below, then write your own solution. The full code is at the bottom.

Approach

Disjoint set union

For Groups of Strings, the implementation maintains connected components and merges them as relationships are processed.

  1. Give each element a component representative.
  2. Merge representatives when a connection is accepted.
  3. Answer connectivity or component queries from the compressed representatives.

Code notes

  • 80 lines of C++ from the credited upstream file 2157.cpp.
  • The implementation visibly relies on sequence storage, hash lookup.
  • 3 loop blocks detected.

Complexity

Account for every find and union operation; with path compression and ranked merging, the amortized cost is nearly constant per operation.

Check the problem constraints before deciding whether this complexity will pass.

Source

Code and credit

This code comes from walkccc LeetCode Solutions by P.-Y. Chen (walkccc) and is used under the MIT licence.

Full codeGroups of Strings · C++C++
Use this to learn the idea, then write your own version.
class UnionFind { public:  UnionFind(int n) : count(n), id(n), sz(n, 1) {    iota(id.begin(), id.end(), 0);  }   void unionBySize(int u, int v) {    const int i = find(u);    const int j = find(v);    if (i == j)      return;    if (sz[i] < sz[j]) {      sz[j] += sz[i];      id[i] = j;    } else {      sz[i] += sz[j];      id[j] = i;    }    --count;  }   int getCount() const {    return count;  }   int getMaxSize() const {    return ranges::max(sz);  }  private:  int count;  vector<int> id;  vector<int> sz;   int find(int u) {    return id[u] == u ? u : id[u] = find(id[u]);  }}; class Solution { public:  vector<int> groupStrings(vector<string>& words) {    UnionFind uf(words.size());    unordered_map<int, int> maskToIndex;    unordered_map<int, int> deletedMaskToIndex;     for (int i = 0; i < words.size(); ++i) {      const int mask = getMask(words[i]);      for (int j = 0; j < 26; ++j)        if (mask >> j & 1) {          // Going to delete this bit.          const int m = mask ^ 1 << j;          if (const auto it = maskToIndex.find(m); it != maskToIndex.cend())            uf.unionBySize(i, it->second);          if (const auto it = deletedMaskToIndex.find(m);              it != deletedMaskToIndex.cend())            uf.unionBySize(i, it->second);          else            deletedMaskToIndex[m] = i;        } else {          // Going to add this bit.          const int m = mask | 1 << j;          if (const auto it = maskToIndex.find(m); it != maskToIndex.cend())            uf.unionBySize(i, it->second);        }      maskToIndex[mask] = i;    }     return {uf.getCount(), uf.getMaxSize()};  }  private:  int getMask(const string& s) {    int mask = 0;    for (const char c : s)      mask |= 1 << c - 'a';    return mask;  }}; 

Did this explanation save you time? I'm a Grade 11 student building this free library to make difficult algorithms easier to understand.

Buy me a coffee ↗